Location: United States
Total raised: $40M
Valuation: $400M
Funding Rounds 1
| Date | Series | Amount | Investors |
| 18.08.2026 | Series A | $40M | a16z |
Mentions in press and media 10
| Date | Title | Description |
| 03.09.2026 | Google’s Gemini 3.8 Flash is built for agents, while its Cyber twin hunts vulnerabilities | Google keeps cranking out Flash models: the company on Wednesday announced two versions of a new 3.8 Flash. The variants include a standard Flash, a “workhorse” model for agentic tasks, software development, and multi-step reasoning, and Fl... |
| 18.08.2026 | Vals AI Raises $40M in Series A Funding at $400M Valuation | Vals AI, a San Francisco, CA-based provider of an artificial intelligence evaluation and model auditing platform, raised $40m in Series A funding at a $400m valuation. The round was led by a16z (Andreessen Horowitz), with participation from... |
| 16.08.2026 | Vals AI Raises $40 Million Series A At $400 Million Valuation As Revenue Grows 8x | Vals AI raised a $40 million Series A at a $400 million valuation as the company scales an independent AI evaluation platform that measures how well frontier models perform on real-world professional tasks. Andreessen Horowitz (a16z) led th... |
| 14.08.2026 | Vals AI Raises $40M From a16z: Frontier Models Fail 52% of Real Finance Analyst Tasks | By Ryan Cook Published: Aug 14 2026, 10:08 AM EDT Share on Facebook Share on Twitter Share on LinkedIn Share on Reddit Share on Flipboard |
| 10.07.2026 | Meta’s Muse Spark 1.1 Opens Paid API at One-Quarter of Anthropic, OpenAI Rates | By Jerry Owens Published: Jul 10 2026, 12:40 PM EDT Share on Facebook Share on Twitter Share on LinkedIn Share on Reddit Share on Flipboard |
| 14.05.2026 | AI IQ is here: a new site scores frontier AI models on the human IQ scale. The results are already dividing tech. | For decades, the IQ test has been one of the most familiar — and most contested — yardsticks for human intelligence. Now, a startup project called AI IQ is applying the same metaphor to artificial intelligence, assigning estimated intellige... |
| 07.03.2026 | GPT-5.4 стал лучшим ИИ для вайб-кодинга | GPT-5.4 занял первое место на Vibe Code Bench v1.1 с результатом 67,42% — на 5,7 п.п. выше предыдущего лидера GPT-5.3 Codex (61,77%). Третье место — у Claude Opus 4.6 без режима рассуждений с 57,57%. Бенчмарк измеряет не умение дописать фун... |
| 28.10.2025 | Anthropic Expands Claude AI for Financial Services | Image: Anthropic Anthropic has made an expansion of its Claude AI platform for financial services, unveiling a suite of new tools designed to bring real-time intelligence and automation to the desks of analysts, bankers, and asset managers.... |
| 28.08.2025 | Nous Research drops Hermes 4 AI models that outperform ChatGPT without content restrictions | Want smarter insights in your inbox? Sign up for our weekly newsletters to get only what matters to enterprise AI, data, and security leaders. Subscribe Now Nous Research, a secretive artificial intelligence startup that has emerged as a le... |
| - | Vals AI | “Private, domain-specific benchmarks in legal, tax, and finance.” |