# Introducing the Open FinLLM Leaderboard

DevFeed: [Introducing the Open FinLLM Leaderboard](<https://devfeed.tech/articles/introducing-the-open-finllm-leaderboard-7317.md>)

Original publisher: [Read original article](<https://huggingface.co/blog/leaderboard-finbench>)

Author: Xie; Jimin Huang; Sophia Ananiadou; Xiao-Yang Liu Yanglet; Alejandro Lopez-Lira; Wang; ldruth; Ruoyu Xiang; chenzhengyu; Yangyang Yu

Published: 2024-10-04T00:00:00Z

Content type: article

Language: en

Sources: [Hugging Face - Blog](<https://devfeed.tech/sources/hugging-face-blog.md>)

Topics: [Finance](<https://devfeed.tech/topics/finance.md>), [LLMs](<https://devfeed.tech/topics/llms.md>), [Human-AI evaluation](<https://devfeed.tech/topics/human-ai-evaluation.md>), [Natural language processing](<https://devfeed.tech/topics/nlp.md>), [Open Source](<https://devfeed.tech/topics/open-source.md>)

Tags: [artificial-intelligence](<https://devfeed.tech/tags/artificial-intelligence.md>), [benchmarks](<https://devfeed.tech/tags/benchmarks.md>), [collaboration](<https://devfeed.tech/tags/collaboration.md>), [community](<https://devfeed.tech/tags/community.md>), [evaluation](<https://devfeed.tech/tags/evaluation.md>), [finance](<https://devfeed.tech/tags/finance.md>), [financial-sector](<https://devfeed.tech/tags/financial-sector.md>), [fine-tuning](<https://devfeed.tech/tags/fine-tuning.md>), [leaderboard](<https://devfeed.tech/tags/leaderboard.md>), [open-source](<https://devfeed.tech/tags/open-source.md>), [qa](<https://devfeed.tech/tags/qa.md>), [testing](<https://devfeed.tech/tags/testing.md>), [text-generation](<https://devfeed.tech/tags/text-generation.md>), [zero-shot](<https://devfeed.tech/tags/zero-shot.md>)

## AI overview

The article introduces the Open FinLLM Leaderboard, a specialized evaluation framework for financial language models. It evaluates models on finance-specific tasks such as information extraction, sentiment analysis, credit risk scoring, stock forecasting, question answering, text generation, and decision-making, using real-world datasets and metrics including Accuracy, F1 Score, ROUGE, and MCC.

## Source excerpt

We're on a journey to advance and democratize artificial intelligence through open source and open science.