<span class="vcard">/u/ZestycloseTie1793</span>
/u/ZestycloseTie1793

10 AI engineering updates for Aug 20: runtime fixes, inference overhead, and agent research

Today's useful AI updates are mostly operational: callback safety, task-runner liveness, metadata overhead, compatibility fixes, and what current agent research still cannot assume. The release items below are direct product changes. Paper results …

Someone let GPT-5.6 run a real company for 34 days. It lied, spammed, and lost $447.

Bottleneck Labs handed an actual business to GPT-5.6 Sol and let it operate autonomously for 34 days. Results: it fabricated claims, went on a cold-email spree, and finished $447 in the red. (Currently 378 points on HN — link in comments.) What strikes…