Six young AI startups now post narrow, measurable wins over established rivals

A scan of recent model launches found six companies founded since 2023 with explicit head-to-head results against established AI products. Liquid AI shows the largest reported edge—3.7× faster in one tool-use test—but all six comparisons are company-reported, so these are scouting signals, not settled market leadership.

6
credible leads to investigate
2023–26
founding years
3.7×
largest reported speed ratio
0
fully independent tests found

Reported uplift over the named rival

Percent improvement; different benchmarks are not directly comparable
blue = company-reported edge
Liquid AI
+270%
Thinking Machines
+66%
Fish Audio
+50%
Poolside
+9.7%
Hark
+5.3%
Perceptron
+0.7%
0%135%270%
Uplift is calculated against the strongest named comparison available, except Thinking Machines, which uses its stated OpenAI quality comparison. Fish Audio is listener-preference lift; Liquid AI is speed lift.

What stands out

1

Hark is the newest and most price-aggressive. Its 2026 browser agent scored 97.7 against GPT‑5.4’s 92.8 while listing token prices far below the named frontier model. venturebeat.com

2

Thinking Machines combines lower latency with higher interaction quality. Its voice/video model reported 0.40-second turn taking versus Gemini’s 0.57 and OpenAI’s 1.18 seconds, plus a 77.8 quality score versus OpenAI’s 46.8. venturebeat.com

3

Liquid AI’s edge is deployment efficiency. Its compact on-device model finished a 35-tool-call workload 3.7× faster than DeepSeek‑V4‑Flash. venturebeat.com

4

The evidence is promising but early. Every shortlisted result originates with the vendor; reproduce the tests on your own workload before treating a benchmark win as durable superiority.

Three strongest scouting targets

best new entrant

Hark

97.7 OM2W · founded 2026

A compelling browser-use score paired with unusually low stated token pricing. Validate reliability on long, messy workflows.

best real-time interaction

Thinking Machines

0.40 sec · founded 2025

The cleanest combined latency-and-quality claim in the set. Access and repeatability remain the key diligence questions.

best edge efficiency

Liquid AI

3.7× faster · founded 2023

Most interesting where cloud dependency, latency, or device constraints matter more than raw general intelligence.

Full shortlist

StartupFoundedProductCompared withReported resultSource
Hark2026HandoffGPT 5.4; Claude Opus 4.897.7 vs 92.8 / 84.1 on OM2W; sharply lower token priceventurebeat.com
Thinking Machines2025TML-Interaction-SmallGemini live; GPT realtime0.40s latency; 77.8 interaction qualityventurebeat.com
Perceptron2024Mk1Robotics‑ER 1.5; Q3.5‑27B85.1 vs 78.4 / about 84.5venturebeat.com
Fish Audio2023S2 ProElevenLabs V360% vs 40% listener preferencefish.audio
Liquid AI2023LFM2.5‑2.6BDeepSeek‑V4‑Flash3.7× faster on a 35-call tool workloadventurebeat.com
Poolside2023Laguna S 2.1DeepSeek‑V4‑Pro‑Max70.2% vs 64.0% on Terminal‑Bench 2.1venturebeat.com
Method: web scan performed 22 August 2026 across recent launch coverage, vendor posts, benchmark comparisons and startup directories; six companies founded in 2023 or later survived the requirement for a named rival and a quantified head-to-head result. “Uplift” measures relative change on each source’s own stated metric and is not a cross-domain ranking.
Keenable · made with SELECT* · 3,772 pages in 2m 22s · Open the chatShare: XLinkedInReddit