## [From Months to Minutes: Rebuilding Our AI Infrastructure for Scale](/content/engineering/rebuilding-ai-infrastructure-scale-llm-gateway/index.html)

WRITER engineers share a deep dive into rebuilding AI infrastructure for scale. Learn how they slashed model integration time from months to minutes. The piece details how a new LLM Gateway replaces hard-coded integrations with a dynamic, self-service platform. Engineers can now add models in seconds, configure guardrails instantly, and track every request with full visibility. Learn how infrastructure was reimagined to serve thousands of agents across hundreds of companies, enabling flexible model choice, real-time supervision, and instant credential rotation—all without downtime.

**Writer Team**

# Writer Engineering

## [How personalized context quietly degrades AI accuracy: a deeper look](/content/engineering/personalized-context-degrades-ai-accuracy/index.html)

**Writer Team**

## [Cerebro: An open source agentic system for security alert triage](/content/engineering/cerebro-ai-security-alert-triage-system/index.html)

**Ben Popper**

## [When too many tools become too much context](/content/engineering/rag-mcp/index.html)

**Ashley Weaver**

## [Rethinking Security: Moving from Human Speed to Machine Speed](/content/engineering/ai-security-machine-speed-defense/index.html)

## [Avoid context rot and improve tool accuracy for AI agents using MCP](/content/engineering/mcp-gateways/index.html)

**Dennis Thompson**

## [Beyond vibe coding: prototyping enterprise agents](/content/engineering/agent-development-lifecycle-prototype/index.html)

**Sam Julien**

## [Palmyra-mini: Small models, big throughput, powerful reasoning](/content/engineering/palmyra-mini-open-source-models/index.html)

**Rakshith Vasudev**

## [WRITER’s Palmyra X5 on Amazon Bedrock: Unlock long context AI for enterprise](/content/engineering/bedrock/index.html)

**Ashley Weaver**

## [The living brain of the enterprise](/content/engineering/enterprise-as-living-brain/index.html)

**Matan-Paul Shetrit**

## [The orchestration graph](/content/engineering/orchestration-graph/index.html)

**Matan-Paul Shetrit**

## [Say hello to Action Agent](/content/engineering/writer-action-agent/index.html)

**Waseem AlShikh**

## [Everyone is a manager now](/content/engineering/employee-into-manager/index.html)

**Matan-Paul Shetrit**

## [Navigating the challenges of generative AI and "vendor lock-in" in enterprises](/content/engineering/vendor-lock-in-generative-ai/index.html)

**Waseem AlShikh**

## [Introducing self-evolving models](/content/engineering/self-evolving-models/index.html)

**Waseem AlShikh**

## [Synthetic data: Busting the myths holding back enterprise AI progress](/content/engineering/synthetic-data-myths-vs-facts/index.html)

**Waseem AlShikh**

## [Short transformers: easily prune redundant LLM layers](https://www.linkedin.com/posts/melisa-russak-5b7987145_short-transformers-easily-prune-redundant-activity-7201220721677635584-lWcg)

**Melisa Russak**
