Try “car wash”, “subscription box”, “Austin” · Esc to close

PhAIL

Benchmark testing real-world performance of AI vision-language models

AI product SaaS & software Show HN · launch post · ▲ 21

Visit site

phail.ai

I built this because I couldn't find honest numbers on how well VLA models [1] actually work on commercial tasks. I come from search ranking at Google where you measure everything, and in robotics nobody seemed to know. PhAIL runs four models (OpenPI/pi0.5, GR00T, ACT, SmolVLA) on bin-to-bin order picking – one of the most common warehouse operations. Same robot (Franka FR3), same objects, hundreds of blind runs. The operator doesn't know which model is running. Best model: 64 UPH. Human teleoperating the same robot: 330. Human by hand: 1,300+. Everything is public – every

Thinking of building something like this?

Every launch here is a competitor to somebody's idea. If yours is close, check it against the market before you build: the Full Check names the rivals, the prices and the gaps.

Check an idea like this

More ai product launches

All

Orion

Visual agent that sees, reasons and acts on images, videos and documents.

AI product SaaS & softwareShow HN ▲ 22

Checked ideas in SaaS & software