I built a fintech agent for discussing financial performance of publicly traded stocks using natural language. I have data covering ~1,500 public companies from scraping I’d already been doing for years.
The AI agent is built with LangChain, Ollama, and Python. It has
- input guardrails
- tool-calls, limits, and tool guardrails
- output guardrails with llm-judge
- context summarization
- set of evals
It uses Meta’s open-weight Llama model, but can be switched to Qwen, Deepseek, Mistral, or about a dozen others instantly.
LLMs are incredible, but they can still fail at the simplest reasoning tasks even when the facts are given to them in a very minimal context history. In this convo using Meta’s Llama model, it says $306B is greater than $391B. So is LLM superintelligence less realistic than a Zuck vs Musk cage fight? And will Zuck challenge Amodei to a cage match next? I’d buy tickets to either of those.