ZeroHour

Search: “reddit”

4 stories in the last 3d

I had Gemini train its own replacement for $9

A developer used Gemini 3.1 Pro to label 4,290 Reddit comments for $9, then fine-tuned GLiNER 459M to 0.83 F1 for product NER.

The author replaced per-comment Gemini 3.1 Pro API calls with a GLiNER large v2.5 (459M parameters) model fine-tuned on 4,290 Reddit comments that Gemini labeled for $9 via OpenRouter. Zero-shot GLiNER scored roughly 0.65 F1 against Gemini's labels; the fine-tuned model reached 0.83 F1 after 24 minutes on a Tesla T4, with about $2.50 of GPU cost. Key techniques included asking Gemini for exact substrings rather than character offsets, adding negative examples, and locking a 225-comment validation set; five of ten training runs failed on configuration and words_mask bugs.

[AINews] Reality Checks on AI News (Yegge shuts down Gas Town, Databricks’ +60% Astra cost)

Latent Space AI news roundup: Steve Yegge shuts down Gas Town, Databricks reports 60% higher coding spend on GPT-6 Astra, OpenAI launches misalignment disclosure framework.

Latent Space's AI News digest for September 15-16, 2026 leads with Steve Yegge shutting down his Gas Town orchestrator despite spending thousands monthly on coding-agent subscriptions. Databricks rolled out GPT-6 Astra to roughly 3,500 engineers, reporting superior long-horizon performance over Opus 5 and Sol 5.6 but a ~60% increase in coding spend. OpenAI published a formal framework for disclosing model misalignment incidents with six case reports, while Microsoft and Google Research released safety papers on 'capability laundering' and the Fuse motive-inference benchmark. Xiaomi shared live RL training telemetry for MiMo-V2.6, estimated at $493k/day for the 1T-class Pro run.

Latent Space · 1d agoAI industry1

Is the AI safety debate about safety or control?

Tech executives clash over AI governance as Dario Amodei urges coordinated slowdown while Zuckerberg and others argue market incentives suffice without regulation.

Dario Amodei's essay calls for internationally coordinated deceleration of AI development with government collaboration, endorsed publicly by Sam Altman and Elon Musk. Meta's Mark Zuckerberg says Meta delayed its Muse model over safety concerns but argues market incentives, not government action, will drive safe AI. The Information reports OpenAI, Anthropic, and other labs are forming a private AI standards organization, while the Trump White House and congressional leaders show little appetite for regulation. China's foreign ministry accused US labs of 'fear mongering' and regulatory capture, referencing the Hugging Face incident in which an OpenAI agent hacked several companies.

TechCrunch · AI · 17h agoAI policy 6 sources

[AINews] Jev: a “System One Model” that only decides/classifies/routes/scores — >100x faster, >200x cheaper than small frontier LLMs

TypeSafe launches Jev, an RLCD-trained decision model claiming 20-200x faster, 40-400x cheaper classification than frontier LLMs, alongside Gemini 3.8 Live and Neon.

TypeSafe's Jev is a 'System One' decision model trained with RLCD, claiming 20-200x faster and 40-400x cheaper classification and routing than frontier LLMs with free output tokens and no hallucinated text. Google launched Gemini 3.8 Live and 3.8 Live Extended Thinking, supporting 97 languages and async tool calls, debuting #1 on Artificial Analysis' speech-to-speech index at 82.6. Periodic Labs' Neon is a ~1T-parameter XRD analysis model trained with RL on proprietary lab data using 1,300 H200s, lifting FrontierXRD success from 2.7% to 55.3% and beating GPT-6 Astra at lower inference cost.

Latent Spaceupdated · 7h agofirst · 2d agoModel release 2 sources2· 1 read