Dispatch
OpenAI Announces $200B Valuation Round   •   EU AI Act Compliance Deadline Extended to 2027   •   Google DeepMind Releases Gemini Ultra 3.0   •   Y Combinator S26 Batch: 60% of Startups Are AI-Native   •   MarTech Consolidation: Salesforce Acquires MadTech Pioneer   •   LLM Token Costs Drop 80% Year-Over-Year   •   Meta Llama 4 Released Under Permissive Commercial Licence   •   Anthropic's Claude Achieves New Benchmarks on Reasoning Tasks   •   Venture Capital Flows to AI Infrastructure Exceed $4B in Q2   •   Adobe GenStudio Reaches 500,000 Enterprise Users   •   OpenAI Announces $200B Valuation Round   •   EU AI Act Compliance Deadline Extended to 2027   •   Google DeepMind Releases Gemini Ultra 3.0   •   Y Combinator S26 Batch: 60% of Startups Are AI-Native   •   MarTech Consolidation: Salesforce Acquires MadTech Pioneer   •   LLM Token Costs Drop 80% Year-Over-Year   •   Meta Llama 4 Released Under Permissive Commercial Licence   •   Anthropic's Claude Achieves New Benchmarks on Reasoning Tasks   •   Venture Capital Flows to AI Infrastructure Exceed $4B in Q2   •   Adobe GenStudio Reaches 500,000 Enterprise Users
Est. MMXXV — Independent Digital PressSunday, 20 September 2026Vol. I — No. 205
MarTech • Startups • LLMs • Digital Strategyterekhindigital.comMorning Edition

Terekhin Digital Media

Rigorous Journalism at the Frontier of Digital Commerce & Machine Intelligence

Sunday, 20 September 2026Issue No. 205
Venture

Vals Raises $40M From a16z to Become the Ratings Agency for AI Models

Series A, 19 September. Confidential test materials prevent exam-gaming that has compromised all major public AI benchmarks. Evaluates across law, finance, coding, cybersecurity, biosecurity, and mental health. Revenue up 8x year-on-year. As Anthropic and OpenAI approach IPOs, third-party model evaluation infrastructure becomes a compliance asset.

Vals raised $40 million led by Andreessen Horowitz on 19 September to scale its confidential AI model evaluation framework — structured as a proprietary benchmarking service in which test materials are never published, preventing the exam-gaming that has compromised all major public AI benchmarks. The company evaluates models across law, finance, coding, cybersecurity, biosecurity, and mental health with task-specific test sets that clients cannot train against. Revenue is up 8x year-on-year; the team has tripled to 25. Co-founder Rayan Krishnan describes the business model as "companies paying to take the SAT" — and the company recently launched a federal agency evaluation track. As Anthropic and OpenAI approach IPOs and AI regulation tightens, third-party independent model evaluation infrastructure is becoming a compliance asset. Vals is positioning to be the authority that certifies model safety and capability claims for the institutional and regulatory audiences that will require independent verification.

ValsAI benchmarkinga16zAI evaluationmodel safetyAndreessen HorowitzAI complianceSeries A
← Return to Front Page
Related Articles
© MMXXVI Terekhin Digital Media — All Rights Reserved — An Independent Digital Publication