Research
Tests of AI agents, retrieval and quant models on real financial data, with the numbers and the method behind every result.
Is Jev efficient for RAG? A test on 10-Ks from Apple, Microsoft, Nvidia and Amazon
A decision model that only returns numbers found the evidence for all 50 analyst questions, with one LLM call per answer and about 100 times fewer tokens than an agentic pipeline.
Sep 27, 2026 · 10 min read

A 12-agent AI investment committee for $0.09. Here's what it decided.
10 AI agents build the case, 2 decide, and neither judge ever reads the data. Inside GitHub's most starred trading project (108,000 stars), tested over 68 runs with every log published.
Sep 24, 2026 · 7 min read

AI finance skills are limited by design. Build your own.
Anthropic ships seventeen finance plugins. Daloopa ships twenty-three skills with 487 stars behind them. Reading the code shows where each one stops, and both packages hardcode the same number.
Sep 21, 2026 · 8 min read

Machine Learning and Deep Learning in Finance: A Complete Guide to 10 Models
Trees, neural networks and real market data, with the code for every one. Three of the ten lost to a baseline that costs nothing, and that result stays in.
Sep 9, 2026 · 15 min read
