Practical AI for Work Efficiency

Clear, grounded write-ups on applying AI to real tasks - research, analysis, productivity, and building things - without the hype.

New Eureka Reports every Wednesday, Saturday and Sunday: AI research, news and inventions, researched by AI agents with every claim sourced. How these are made

The 60-second version

  • Google's Gemini 4 Argon, announced Sept. 30, scored 53 on Artificial Analysis's Intelligence Index to match GPT-6 Astra and ranked first on the Vals Index at 68.90%, while several of Google's own headline scores were computed by Google itself.
  • A preprint not yet peer-reviewed, by 15 authors including researchers at Rutgers and McGill (most list themselves as independent researchers), found that self-training Qwen3.5 search agents agreed on the same wrong answer up to 8.8% of the time by round three, and a fix called CrossFit cut that to 3.7% while raising benchmark scores by about 8 points.
  • Stillwet, a project by its creator Alice, shows 75 oil paintings made by AI models writing code for a paint simulator, and in one small blind test, three AI judges each ranked their own painting 5th or 6th of six.

Becoming Irreplaceable With AI

Read more articles

Recently published

See all articles

New articles on applying AI, straight to your inbox

No noise, just practical write-ups on how to put AI to work.