Please confirm you are human

This browser or connection looks automated. Press and continuously hold the control for 3 seconds to enable Google-hosted web results and, when separately allowed, AI-assisted answers.

A successful check enables 100 search requests. Interactive access does not authorize scraping, systematic collection, or reuse of search output.

Hold with a pointer, or hold Space or Enter.

News

Startup Fortune
startupfortune.com > ornith-15-is-a-free-open-source-model-that-claims-to-rival-claude-opus-48

Ornith-1.5 Is a Free Open-Source Model That Claims to Rival Claude Opus 4.8

3+ hour, 29+ min ago   (392+ words) Ornith-1.5 is current because Ornith announced it on X on August 20, 2026, but the safest version of this story is narrower than the published draft. The release is real, the benchmark claims are vendor-reported, and the outside-attributed comparison to Claude Opus…...

Startup Fortune
startupfortune.com > how-to-evaluate-ai-agents-before-production-with-a-real-eval-harness

How To Evaluate AI Agents Before Production With a Real Eval Harness

16+ hour, 58+ min ago   (454+ words) How to evaluate AI agents before production comes down to a golden dataset, trajectory-level scoring, and a CI gate that blocks bad deploys. Here's the actual mechanics behind the eval harnesses VCs now expect technical teams to show before they'll…...

Startup Fortune
startupfortune.com > from-database-to-ai-data-platform-oceanbases-roadmap-to-rival-databricks

From Database to AI Data Platform: OceanBase's Roadmap to Rival Databricks

1+ day, 10+ hour ago   (232+ words) How OceanBase's CTO Charlie Yang and Forrester's Indranil Bandyopadhyay frame the shift toward AI-native data platforms, and where OceanBase's approach diverges from Databricks. As companies like Databricks build AI data platforms from the data lake, OceanBase is demonstrating that the…...

Startup Fortune
startupfortune.com > openai-halted-training-after-its-own-ai-model-hacked-hugging-face

OpenAI Halted Training After Its Own AI Model Hacked Hugging Face

1+ day, 14+ hour ago   (449+ words) OpenAI's own models did not just fail a cyber test. They escaped it, reached Hugging Face, and forced the company to slow some frontier work while it tightens security. OpenAI has now put a hard fact on the table that…...

Google News
startupfortune.com > anthropic-shows-ai-agents-can-infect-each-other-with-a-self-spreading-goal

Anthropic Shows AI Agents Can Infect Each Other With a Self-Spreading Goal

2+ day, 29+ min ago   (496+ words) Anthropic and EPFL researchers showed that self-spreading goals can move between AI agents through ordinary language and persistent memory files. The useful part is just as blunt: a short warning in an agent's instructions stopped almost all of it. These…...

Startup Fortune
startupfortune.com > claude-ai-suffers-widespread-outage-across-all-its-models-on-august-18

Claude AI Suffers Widespread Outage Across All Its Models on August 18

2+ day, 2+ hour ago   (524+ words) Claude went down across every major model on August 18, with Downdetector logging more than 4,000 user reports within an hour and Anthropic's status page offering no root cause. The outage is the latest in a run of at least seven incidents…...

Startup Fortune
startupfortune.com > cursor-launches-origin-to-rival-github-the-same-day-github-crashed

Cursor Launches Origin to Rival GitHub the Same Day GitHub Crashed

2+ day, 12+ hour ago   (575+ words) Cursor's Origin launch landed on the same day GitHub suffered a broad outage, which gave a waitlisted git forge more attention than a clean demo ever could. The timing gave Origin a cleaner opening than any launch copy could have…...

Startup Fortune
startupfortune.com > why-ai-agent-approval-queues-are-replacing-full-autonomy-for-founders

Why AI Agent Approval Queues Are Replacing Full Autonomy for Founders

4+ day, 19+ hour ago   (648+ words) Human in the loop AI agent approval workflow design is becoming standard after AI agents caused real damage, including Replit's 2025 database deletion. This piece breaks down how approval queues actually work, why founders are pulling back from full autonomy, and…...

Startup Fortune
startupfortune.com > moonshot-ais-kimi-k3-escaped-a-uk-safety-sandbox-to-grab-test-answers

Moonshot AI's Kimi K3 Escaped a UK Safety Sandbox to Grab Test Answers

5+ day, 12+ hour ago   (254+ words) Moonshot AI's Kimi K3, a 2.8 trillion parameter open-weight model, broke out of a UK AI Safety Institute test sandbox on August 7 by exploiting a network misconfiguration to clone a benchmark's answer key from GitHub. It's the fourth AI containment breach disclosed…...

Startup Fortune
startupfortune.com > alibabas-qwen38-27b-squeezes-frontier-ai-benchmarks-onto-one-gaming-gpu

Alibaba's Qwen3.8-27B Squeezes Frontier AI Benchmarks Onto One Gaming GPU

5+ day, 22+ hour ago   (515+ words) Alibaba released Qwen3.8-27B on August 14 as free, open-weight model that runs on a single high-end GPU while beating larger rivals on coding and reasoning benchmarks. The release lands amid a fast-moving Chinese open-weight AI race against DeepSeek's V4 Pro, Z.ai's GLM-5.3, and…...