WRITER Engineering blog | AI development news
From Months to Minutes: Rebuilding Our AI Infrastructure for Scale
WRITER engineers share a deep dive into rebuilding AI infrastructure for scale. Learn how they slashed model integration time from months to minutes. The piece details how a new LLM Gateway replaces hard-coded integrations with a dynamic, self-service platform. Engineers can now add models in seconds, configure guardrails instantly, and track every request with full visibility. Learn how infrastructure was reimagined to serve thousands of agents across hundreds of companies, enabling flexible model choice, real-time supervision, and instant credential rotation—all without downtime.
Writer Team
Writer Engineering
How personalized context quietly degrades AI accuracy: a deeper look
Writer Team
Cerebro: An open source agentic system for security alert triage
Ben Popper
When too many tools become too much context
Ashley Weaver
Rethinking Security: Moving from Human Speed to Machine Speed
Avoid context rot and improve tool accuracy for AI agents using MCP
Dennis Thompson
Beyond vibe coding: prototyping enterprise agents
Sam Julien
Palmyra-mini: Small models, big throughput, powerful reasoning
Rakshith Vasudev
WRITER’s Palmyra X5 on Amazon Bedrock: Unlock long context AI for enterprise
Ashley Weaver
The living brain of the enterprise
Matan-Paul Shetrit
The orchestration graph
Matan-Paul Shetrit
Say hello to Action Agent
Waseem AlShikh
Everyone is a manager now
Matan-Paul Shetrit
Navigating the challenges of generative AI and "vendor lock-in" in enterprises
Waseem AlShikh
Introducing self-evolving models
Waseem AlShikh
Synthetic data: Busting the myths holding back enterprise AI progress
Waseem AlShikh
Short transformers: easily prune redundant LLM layers
Melisa Russak