Archive: 2026/07
Browse posts by year and month
2026
28 postsJuly 28 posts
- LLM-Native Recommendation Architecture After Netflix GenRec
- How MCP Servers for AI Agents Bridge Fragmented Data
- Control Reasoning Effort LLM APIs in Production
- BM25 vs Dense Retrieval and SPLADE for Production RAG
- Shipping LLM Browser Agents Without Breaking Production
- Claude Opus 5 Prompt Injection Hits 0% Across 129 Tests
- AI Coding Agents Cut Per-Task Cost Up to 4x With Indexing
- ChatGPT Shared Link Vulnerability Plants Rogue Agents
- Grok vs Copilot in Excel for Builders
- Run AI Models Locally on Mac With MLX and Nativ
- Hugging Face Agentic Attack Redefines AI Security Response
- AI Agent Cost Per Resolution Decides If It Ships
- Self-Host LLM Inference, the Netflix Decision Framework
- Where AI in Media Production Workflows Actually Pays Off
- Orchestrate Open Source LLMs vs Frontier Models
- OpenAI Codex Subagent Encryption Breaks Agent Observability
- ChatGPT Sales Workflows Fail on Real CRM Data
- Real-Time AI Dental Image Verification Cuts Claim Denials
- Fine-Tuning vs RAG vs Prompt Engineering Decision Framework
- How LLM Agent Scaffolding Fixes Failing Code Review Agents
- Building Proactive AI Agents With Context Graphs
- AI Coding Benchmarks Are Gameable by Design
- AI Video Editor Limitations Break Iterative Workflows
- LLM Vendor Data Risk Has a Break-Even Price
- Stop RAG Hallucination with Typed Schema Contracts
- Domain-Specific LLM Evaluation Demands Leakage-Free Data
- Why an Agent-to-Agent Gateway Beats Point-to-Point Links
- LLM API Token Inflation Hides Claude Sonnet's Real Cost