Cover Image for Building Trustworthy Financial AI: Governance, Bias, and Mechanistic Insights
Cover Image for Building Trustworthy Financial AI: Governance, Bias, and Mechanistic Insights
Avatar for Data Phoenix
Presented by
Data Phoenix

Building Trustworthy Financial AI: Governance, Bias, and Mechanistic Insights

Virtual
Registration
Welcome! To join the event, please register below.
About Event

Large Language Models are increasingly being integrated into financial workflows, supporting tasks ranging from investment analysis to decision-making assistance. However, these systems can exhibit subtle biases that may impact the quality and fairness of their recommendations. In this talk, I will present our research on positional bias in financial LLMs, a phenomenon where model decisions are influenced by the order in which information is presented rather than the information itself. I will also discuss how mechanistic interpretability techniques can be used to identify the internal circuits responsible for these behaviors and how these findings contribute to AI governance frameworks for financial applications.

Key Highlights

  • Understanding positional bias in financial AI systems

  • Measuring bias across open-source large language models

  • Using mechanistic interpretability to trace model behavior

  • Implications for AI governance, risk management, and regulatory compliance

  • Best practices for building more trustworthy financial AI systems


Speakers

Fabrizio Dimino is an AI Research Scientist at Domyn specializing in trustworthy AI for financial services, with experience building agentic systems and predictive ML models. Published research at ICAIF, NeurIPS GenAI in Finance, ICLR FinAI, IEEE ICDM on knowledge graphs, AI governance, red-teaming, and reinforcement learning. His recent research investigates how biases emerge in large language models and how these systems can be made more reliable for high-stakes financial applications.

Launch partner

AgentField is open-source infrastructure for building autonomous software factories and the AI backends that power them. It gives multi-agent systems a single control plane for orchestration, governance, and provenance - so every action your agents take is policy-checked and accountable, with no glue code and no editable logs. Apache 2.0, with SDKs in Python, TypeScript, and Go.

Partners

AI Tabir is an AI adoption platform that helps small businesses run like self-driving companies — AI takes on the operational busywork while owners stay in control of the decisions that matter. It starts with a free assessment that maps how a business actually works and pinpoints where time and money leak out — missed calls, unanswered leads, quotes with no follow-up, overdue invoices — then deploys industry-tuned AI agents, managed end to end, to handle that work. Rather than selling another tool, the platform delivers the work itself and keeps adapting as each business grows, across industries including automotive, HVAC, media, and video production.

KROK

The AI Collective is a global non-profit building the human layer for the AI era. We unite 200,000+ leaders, builders, and stakeholders across 150+ forums worldwide to democratize the frontier, build trust, and coordinate how society navigates the rapid acceleration of technological progress.

Avatar for Data Phoenix
Presented by
Data Phoenix