Archive: 2026/08
Browse posts by year and month
2026
19 postsAugust 19 posts
- AI Agent Memory Lessons From LinkedIn's Hiring Assistant
- Why Speculative Decoding Pays Nearly 4x on CPUs
- AI Agent Token Usage Overtook Humans on OpenRouter
- How WhatsApp Scam Alert Detects Scams It Cannot Read
- AI-Assisted Product Launch Teardown of Stampli's 68% Claim
- Build an AI Text Detector, Then Test If It Can Ship
- The Real Cost of Gradient Accumulation on T4 and L4
- LLM Context Window Management With a Token Budget
- The Real OpenAI Ultrafast Mode Speedup, Workload by Workload
- LangChain vs LangGraph for Stateful AI Agent Orchestration
- ChatGPT Business Premium Pricing Decodes Agent Token Math
- Muse Glimmer Local Review Tests 30B Agents Under 20GB
- Structured Output Local LLM Tactics That Survive Production
- AI Agent Cyber Security Evaluation After the Astra Slowdown
- Agentic Loop Token Costs Are an Architecture Problem
- Megakernels in LLM Inference When Fusion Actually Wins
- Stacked Pull Requests Relocate AI Mega-PR Review Cost
- Disaggregated GPU Inference Hits the KV Cache Wall
- How The Copilot Prompt Injection Worm Spreads