See Aiera in Action
Whether you’re looking for a quick walkthrough or enterprise pricing, our team is ready to show you how Aiera can fit your needs.
Whether you’re looking for a quick walkthrough or enterprise pricing, our team is ready to show you how Aiera can fit your needs.
We may request cookies to be set on your device. We use cookies to let us know when you visit our websites, how you interact with us, to enrich your user experience, and to customize your relationship with our website.
Click on the different category headings to find out more. You can also change some of your preferences. Note that blocking some types of cookies may impact your experience on our websites and the services we are able to offer.
These cookies are strictly necessary to provide you with services available through our website and to use some of its features.
Because these cookies are strictly necessary to deliver the website, refusing them will have impact how our site functions. You always can block or delete cookies by changing your browser settings and force blocking all cookies on this website. But this will always prompt you to accept/refuse cookies when revisiting our site.
We fully respect if you want to refuse cookies but to avoid asking you again and again kindly allow us to store a cookie for that. You are free to opt out any time or opt in for other cookies to get a better experience. If you refuse cookies we will remove all set cookies in our domain.
We provide you with a list of stored cookies on your computer in our domain so you can check what we stored. Due to security reasons we are not able to show or modify cookies from other domains. You can check these in your browser security settings.
We also use different external services like Google Webfonts, Google Maps, and external Video providers. Since these providers may collect personal data like your IP address we allow you to block them here. Please be aware that this might heavily reduce the functionality and appearance of our site. Changes will take effect once you reload the page.
Google Webfont Settings:
Google Map Settings:
Google reCaptcha Settings:
Vimeo and Youtube video embeds:
You can read about our cookies and privacy settings in detail on our Privacy Policy Page.
Privacy Policy
Triangulating Value with AI in Research
The Aiera Research Score ranks the response quality of 32 language models on investment research queries. But fold in what each model costs (price, speed, and tokens per task) and the #1 model drops to 29th.
Our Research Score measures one thing: how well does a model produce institutional-grade analysis, across Facts, Grounding, and Depth. At the moment, Opus 5 tops the leaderboard with a score of 87.
That score represents the model’s capability. But we also need to consider how to rank a model’s value.
Three numbers no capability board can show decide value:
We combined all of these factors into a single Value Score: research quality, average cost per task, and average latency per task. Here’s what happens to the board:
How the Value Score works
Performance is the Aiera Research Score (mean of Facts, Grounding, Depth). We measured each model’s average output tokens per task (and time to completion) across all tasks included in the benchmark, then turned price and speed into what you actually experience:
cost / task = tokens × output price · latency / task = tokens ÷ speed
Value = 0.5·performance + 0.3·cost + 0.2·latency (each normalized 0–100)
Performance carries the most weight on purpose; a cheap, fast model that can’t do the work is worthless. Price and speed then break the ties a capability board ignores.
What the Token Count Reveals
The most capable models are also the most verbose. Opus 5 and Sonnet 4.6 write ~5,900 tokens per answer — more than double the non-Anthropic model . At premium output prices, that length compounds twice: it’s what you’re billed for and what you wait through. Opus 5’s brilliance arrives at ~15¢ and 115 seconds per task. A per-token price tag hides this entirely; a per-task view makes it the headline.
The value board is led by models that are both strong in research and economical with words. MiniMax M3 (81.5 research) and DeepSeek V4 Pro (80.1) deliver near-frontier quality research output for well under a cent per task in less than half the time; Opus 5 costs roughly 30–50× more per answer for a few points more quality. Even inside OpenAI’s own lineup the split is stark: the efficient GPT-5.6 Luna lands #2 on value, while the flagship GPT-5.6 Sol (nearly tied on research) sits at #22, sunk by a wordy 4,100-token task at premium frontier prices.
Time-to-answer is where verbosity and throughput collide. Nemotron 3 Ultra, Gemini 3.5 Flash, and Gemini 3.1 Pro return answers in under 15 seconds because they’re both fast and concise. The Kimi models are strong and cheap but take over a minute; not from weakness, but because they’re slower and wordier.
The picks
Best overall value: MiniMax M3
81.5 research (top-10 quality) at 0.45¢ and ~49s per task. Near-frontier work at open-weight economics.
Frontier-grade, no tax: GPT-5.6 Terra or DeepSeek V4 Pro
80+ research at a fraction of the flagship cost and half the latency; the sweet spot when quality can’t slip.
Fastest capable: Nemotron 3 Ultra or Gemini 3.5 Flash
70/68 research with answers in 10–13 seconds. When a person is waiting, speed is the feature.
The Net Net
Value needs a quality floor. Rank purely on cost and the cheapest, fastest models like Grok 4.3, Llama 4 Scout, and gemma would top the chart. They don’t, because they score under 30 on research: dividing quality you don’t have by a price you barely pay is a useless recommendation. That’s why performance carries half the weight, and why those models sit low despite near-zero cost.
Opus 5 really is the best on research quality. If your task is rare, high-stakes, and latency doesn’t matter (a one-off memo, not a pipeline) pay for the ceiling.
The value board isn’t an argument against the best model. It’s a map of what the best model costs, and of how little you give up to run something cheaper and / or faster at scale. There’s no single best model, only a best model for a budget, a latency target, and a quality floor.
Explore the Aiera Leaderboard >