Observed Signal · Jan 22, 2026 · Product & Technology Development · Source: OpenAI Blog · Impact: 3/5 · Sentiment: Positive

Powering 800 Million ChatGPT Users with PostgreSQL

Executive Signal Summary

OpenAI published an engineering post describing how it scaled PostgreSQL to support millions of queries per second and serve 800 million ChatGPT users. The team retained a single-primary Azure PostgreSQL flexible server for writes and operated nearly 50 geo-distributed read replicas, while migrating shardable, write-heavy workloads to sharded systems such as Azure Cosmos DB. Key techniques included aggressive query optimization, workload isolation, PgBouncer connection pooling (reducing average connection time from 50ms to 5ms), cache locking to prevent cache-miss storms, rate limiting, strict schema-change controls, and testing cascading replication with Azure to scale replicas. OpenAI reports low p99 read latency, five-nines availability, and only one SEV-0 Postgres incident in the past 12 months.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

OpenAI demonstrates that PostgreSQL can be scaled to very large read-heavy workloads using cloud-managed Postgres, read-replica strategies, connection pooling, and migration of writes to sharded systems. The techniques and operational lessons are relevant to large-scale platform and data infrastructure teams.

SIGNAL RADAR

Track Microsoft Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • OpenAI’s PostgreSQL load grew more than 10x over the past year.
  • OpenAI uses a single-primary Azure PostgreSQL flexible server instance with nearly 50 read replicas across regions.
  • OpenAI migrated shardable, write-heavy workloads to sharded systems such as Azure Cosmos DB and disallows adding new tables to the PostgreSQL deployment.
  • OpenAI deployed PgBouncer for connection pooling, reducing average connection time from ~50 ms to ~5 ms in benchmarks.
  • OpenAI is testing cascading replication with Azure PostgreSQL to scale beyond ~50 replicas while managing WAL shipping and failover complexity.
Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: OpenAI Blog•Published: Jan 22, 2026
Original Coverage Title: “Scaling PostgreSQL to power 800 million ChatGPT users”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

InfrastructureSep 11, 2026

OpenAI Scales Habitat Storage to 1 Billion Users

OpenAI has shared details on scaling its internal storage platform, Habitat, to support over 1 billion weekly users. Habitat now handles more than 70 million requests per second and serves over 500 petabytes of data across nearly 40 regions. The platform evolved from a simple Python client-side library into a standalone service to handle the operational complexity of coordinating deployments across multiple products. Key decisions included adopting Python for flexibility despite performance trade-offs, then rewriting the service in Rust for efficiency. The engineering team addressed issues like asyncio scheduling delays, metastable failures due to LIFO connection pooling, and thundering herd problems by using Envoy for connection management. Habitat uses Azure Cosmos DB for storage and offers an offline analytical view via Rockset. The post is part one of a series on scaling online storage infrastructure.

Read assessment
AI & AutomationSep 11, 2026

OpenAI Launches Data Agent in ChatGPT Work, Halts Pro Subscriptions on Astra Demand

OpenAI has introduced a new Data Agent feature within ChatGPT Work, allowing users to query company data using natural language, perform root-cause analysis, and create interactive dashboards without needing SQL skills. The agent connects to various data sources including Amazon Redshift, Google BigQuery, Databricks, MongoDB, Snowflake, ClickHouse, and Datadog, as well as files from Google Drive and SharePoint. It respects existing access controls and can incorporate internal definitions and relationships from tools like dbt, Snowflake Horizon, and Databricks Genie Ontology. Additionally, OpenAI is halting new Pro subscriptions due to unprecedented demand for its GPT-6 Astra model, which was launched in early September 2026. The company cites the high computational load from Pro users as the reason. Existing Pro users are unaffected, and other tiers remain available.

Read assessment
Enterprise AI / Data AnalyticsSep 10, 2026

OpenAI Launches Data Agent for ChatGPT Work

OpenAI introduced a new Data agent within ChatGPT Work, designed to let business users analyze company data through natural language queries. The agent connects to various data sources including Amazon Redshift, Datadog, Google BigQuery, ClickHouse, Databricks, MongoDB, and Snowflake, as well as files from Google Drive and SharePoint. It leverages semantic layers and business context from platforms like Databricks Genie, dbt, GitHub, and Snowflake Horizon. Users can build and share interactive dashboards, and the agent can also interact with BI tools like Tableau, Power BI, Sigma, and ThoughtSpot. OpenAI states that the agent is built from internal tools, with most of its product team and over two-thirds of its GTM organization using it. Early adopters include NTT Data, Thermo Fisher, ServiceTitan, and others.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.