Zero-Shot LLM Evaluation for MiFID II Trading Reports Using Chain-of-Thought Prompting

Authors

  • Kalyan Kondisetty Wavicle Data Solutions, USA Author
  • Tejas Dhanorkar Discover Financial Services, USA Author

Keywords:

MiFID II, large language models, zero-shot evaluation, chain-of-thought prompting, regulatory compliance

Abstract

This research uses chain-of-thought (CoT) prompting to analyze zero-shot large language model (LLM) evaluation techniques for MiFID II trading reports to increase reasoning transparency and regulatory compliance verification. To evaluate their MiFID II-compliant discrepancies, omissions, and inconsistencies detection, GPT-4 and Gemini algorithms make autonomous logical judgments on trade reporting datasets. To assess model resilience, the research thoroughly tests these LLMs' reasoning trails for coherence, compliance accuracy, and robustness under adversarial testing scenarios using purposely confused or faked trade data. CoT-enabled zero-shot automated regulatory supervision systems' operational correctness strengths and shortcomings are compared using expert-authored validation checklists. LLMs may increase compliance, remove manual review, update transaction reporting, and discover model interpretability and dependability issues. AI-driven reasoning regulatory technology frameworks for financial market compliance automation may benefit from these discoveries.

Downloads

Download data is not yet available.

References

E. Ferran, “MiFID II: The challenge of implementation,” Journal of Financial Regulation and Compliance, vol. 26, no. 3, pp. 230–249, 2018.

European Securities and Markets Authority (ESMA), “MiFID II and MiFIR: Regulatory Technical and Implementing Standards,” Official Journal of the European Union, 2018.

A. Arner, J. Barberis, and D. W. Buckley, “RegTech: Transforming financial regulation through technology,” Journal of Financial Perspectives, vol. 3, no. 3, pp. 1–14, 2016.

T. Philippon, “The FinTech opportunity,” NBER Working Paper No. 22476, 2016.

D. R. Bell and C. H. O’Reilly, “Artificial intelligence in compliance: Risks, challenges, and opportunities,” Harvard Business Review on Financial Regulation, vol. 5, pp. 44–59, 2020.

J. Brown, M. Sutherland, and E. Kominers, “Automated compliance in financial markets: Machine learning approaches under MiFID II,” European Financial Management Journal, vol. 28, no. 2, pp. 315–337, 2022.

J. Wei, X. Wang, D. Schuurmans, Q. Le, and D. Zhou, “Chain-of-Thought Prompting Elicits Reasoning in Large Language Models,” Advances in Neural Information Processing Systems (NeurIPS), vol. 35, pp. 24824–24837, 2022.

T. Kojima, S. Gu, M. Reid, Y. Matsuo, and Y. Iwasawa, “Large Language Models are Zero-Shot Reasoners,” arXiv preprint arXiv:2205.11916, 2022.

J. OpenAI et al., “GPT-4 Technical Report,” arXiv preprint arXiv:2303.08774, 2023.

Google DeepMind, “Gemini 1 Technical Overview,” Google Research Technical Report, 2024.

S. Chia, Y. Tay, J. Su, D. Bahri, and D. Metzler, “Can large language models reason about regulations? Benchmarking interpretability and compliance logic,” Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (ACL), pp. 405–421, 2023.

N. Mishra, D. Sahoo, and S. Hoi, “Automated compliance validation using machine learning: A RegTech perspective,” ACM Transactions on Management Information Systems, vol. 13, no. 4, pp. 1–25, 2022.

M. Srivastava and H. Chockler, “Explainability and reasoning in AI-based compliance systems,” AI & Society, vol. 38, pp. 799–816, 2023.

A. Bommasani et al., “On the opportunities and risks of foundation models,” Stanford Center for Research on Foundation Models (CRFM) Technical Report, 2021.

P. Thiebes, M. Lins, and A. Sunyaev, “Trustworthy artificial intelligence,” Electronic Markets, vol. 31, no. 2, pp. 447–464, 2021.

M. Xie, S. Liu, and E. Cambria, “Explainable AI for financial regulation: Towards auditable and transparent decision systems,” IEEE Transactions on Computational Social Systems, vol. 9, no. 1, pp. 64–78, 2022.

R. Bommarito, D. Katz, and M. Zelner, “Regulatory machine learning: Computational models for automated legal compliance,” Artificial Intelligence and Law, vol. 29, no. 3, pp. 377–403, 2021.

European Parliament and Council, “Regulation (EU) 2024/1680 on Artificial Intelligence (AI Act),” Official Journal of the European Union, 2024.

C. Molnar, Interpretable Machine Learning: A Guide for Making Black Box Models Explainable, 2nd ed., Leanpub, 2022.

J. Zhao, A. Liu, and T. Zhang, “Evaluating robustness and bias in large language models for compliance applications,” Proceedings of the IEEE International Conference on Trust, Security and Privacy in Computing and Communications (TrustCom), pp. 222–233, 2023.

Downloads

Published

20-12-2022

How to Cite

[1]
Kalyan Kondisetty and Tejas Dhanorkar, “Zero-Shot LLM Evaluation for MiFID II Trading Reports Using Chain-of-Thought Prompting ”, Art. Intel. Mach. Learn. Auto. Sys., vol. 6, pp. 116–149, Dec. 2022, Accessed: Jul. 28, 2026. [Online]. Available: https://amlas.net/index.php/publication/article/view/48

Similar Articles

1-10 of 47

You may also start an advanced similarity search for this article.