Scale
We track 60 posts about Scale from 40 engineering blogs. Most active: Nvidia, Facebook, Thumbtack. Latest post: Oct 7, 2026.
Companies writing about Scale
Recent posts
Workbench: Building an AI Content Generation and Quality Assurance System at Scale (opens on the source site)
How Thumbtack built an AI pipeline that generates, evaluates, and refines marketing content at scale, maintaining human-level quality programmatically.The ChallengeThumbtack connects customers with local service professionals across a wide range of categories and geographies. For many customers, the first entry point to Thumbtack is a landing page. When someone searches Google for “plumber near me” or “duct cleaning Raleigh NC”, these pages are where they land, and they are the user’s first impression of the marketplace for that query. The footprint targeted by this work is roughly 500K such…
Secret protection must scale with software (opens on the source site)
Developers aren’t becoming more careless; they’re being outpaced. The tools that let developers create more software should also take on more of the work of protecting it. The post Secret protection must scale with software appeared first on The GitHub Blog.
Building Git infrastructure for agent-scale development (opens on the source site)
Writes are accelerating, and this growth can add stress to systems all repos depend on. Here's our approach to building architecture that can scale. The post Building Git infrastructure for agent-scale development appeared first on The GitHub Blog.
NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI (opens on the source site)
Nvidia ·
Local AI is becoming more useful by the token. As AI agents move from experiments into everyday development, increasingly capable open models are shrinking to fit on more devices, giving builders more to run locally. Coming this month, NVIDIA DGX Spark will be available with 64GB of unified memory from top manufacturer partners — Acer, […]
Automated Database Failover: Why Homegrown High Availability Struggles At Scale (opens on the source site)
I’ve encountered DBAs and ops engineers who shared war stories of running production long before automation and modern tooling went mainstream. Experiencing sudden primary faults during sleeping hours, getting paged, logging in to squint over trace logs that show replication lag on all replicas and working out which one is furthest ahead, promoting it, editing […] The post Automated Database Failover: Why Homegrown High Availability Struggles At Scale appeared first on Severalnines.
GPU virtualization at scale: AMD MI300X SR-IOV on Red Hat OpenStack Services on OpenShift (opens on the source site)
Red Hat ·
Enterprise AI environments are increasingly moving from dedicated GPU servers toward shared accelerator infrastructure that can support multiple users, teams, and workloads on the same hardware. This shift creates an important infrastructure question: How can organizations increase GPU utilization while maintaining strong workload isolation and predictable performance? The post GPU virtualization at scale: AMD MI300X SR-IOV on Red Hat OpenStack Services on OpenShift appeared first on Red Hat Developer.
Prioritize your messages at scale with the Twilio Traffic Optimization Engine (opens on the source site)
Twilio ·
The Traffic Optimization Engine is a suite of products working across networks, senders, geography, and use cases to provide messaging deliverability at scale.
Cloudflare Containers, rebuilt to scale agent sandboxes (opens on the source site)
Cloudflare Containers now start 6x faster, let your agent choose each sandbox's image and instance type at runtime, and support filesystem snapshots in public beta, all controlled from a Durable Object.
Sakeena Fiza Helps NVIDIA Hardware Succeed at Scale (opens on the source site)
Nvidia ·
When Sakeena Fiza describes her work as a validation engineer at NVIDIA, she does so in terms more befitting a detective story than a world-class engineering lab. “Validation engineers look in the shadows and shine a light into every corner,” Fiza said. “Every time we get a system, our first thought is: how can it […]
What Dropbox has learned from deploying AI at company scale (opens on the source site)
Dropbox ·
Learnings from deploying AI at company scale, and what organizations need to consider as they move from AI adoption to broader transformation.
Why Deploying Physical AI at Scale Demands Safety at Every Layer (opens on the source site)
Nvidia ·
Physical AI is moving rapidly from research to large-scale deployment. By 2035, ABI Research projects an installed base of 49 million level 3-5 autonomous vehicles (AVs), while Omdia estimates that roughly 60 million industrial robots will be deployed between 2026 and 2035. As these machines enter roads, factories, warehouses and other environments shared with people, […]
From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale (opens on the source site)
Nvidia ·
Today, Egypt’s AI builders gathered in the Grand Egyptian Museum for a reception that highlighted the nation’s rapidly growing AI ecosystem — spanning AI natives, developers, researchers, startups and enterprises — building applications across industries. The event included a keynote from Paolo Guglielmini, vice president of EMEA at NVIDIA. Ahmed Mostafa, regional AI adoption lead […]
d-Matrix Adopts NVIDIA NVLink Fusion for Rack-Scale XPU Deployment (opens on the source site)
Nvidia ·
AI inference chipmaker d-Matrix today announced it will use NVIDIA NVLink Fusion to connect its next-generation Raptor XPUs to NVIDIA’s AI infrastructure platform — joining a growing roster of ecosystem partners. By connecting Raptor to NVIDIA NVLink scale-up and Spectrum-X scale-out networking, the NVIDIA MGX rack architecture and the broader NVIDIA AI platform, NVLink Fusion […]
Running Playwright at scale: connecting to the Zyte CDP browser (opens on the source site)
Running Playwright at scale: connecting to the Zyte CDP browserBlogRunning Playwright at scale: connecting to the Zyte CDP browserArticleTutorial / How-to Running Playwright at scale: connecting to the Zyte CDP browser John Rooney · Developer Engagement Manager September 7, 2026 Try Zyte API Build your first scraper in minutes Free trial, no credit card. From a single request to production in an afternoon. Get started John Rooney Developer Engagement Manager John is the Developer Engagement Manager at Zyte, working closely with the community, creating content and helping developers learn web…
The economics of agent scale: tokens, ROI, and building platforms for AI-first teams (Part 2) (opens on the source site)
Andi Gutmans, head of Agentic Data Cloud at Google, returns for the second half of his Leaders of Code conversation to talk through the cost and infrastructure side of agentic development. ICYMI, part one covered judgment, code review, and data activation.
MAPS: Netflix’s Multimodal Asset Personalization at Scale (opens on the source site)
Netflix ·
By Emma Yanyang Kong, Aditya Deshpande, Asad Abbasi, Bowei Yan, David Fagnan, Ashish Rastogi, Dhaval Patel, Ray ZhangIntroductionThe Netflix experience is a journey of discovery. Every visual cue, from the artwork on a title to the video previews that autoplay while you browse, is there to connect you with a story you will love. We call these visual cues assets, and choosing the right one for each member is a personalization problem of its own. But which image or video preview of Squid Game should we show you? And what do we do right after a title launches, when there’s far too little…
MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet (opens on the source site)
Facebook ·
Training and serving frontier AI models depends on fast, reliable networks that move data between GPUs without wasting compute cycles. To meet this challenge at scale, Meta designed MetaRoCE – a clean-sheet RDMA transport protocol purpose-built for AI workloads on commodity Ethernet. We’re releasing the MetaRoCE specification, a reference software implementation and a compliance test [...] Read More... The post MetaRoCE: A New RDMA Transport Built for AI-Scale Ethernet appeared first on Engineering at Meta.
Harnessing the agent semantic reliability at scale (opens on the source site)
The engine can produce both inferential and deterministic knowledge. A relation the engine surfaces but hasn't confirmed stays inferential, an open judgment call, until a domain expert reviews it. Once approved, the ontology, the relation rules and the ownership record behind each one all become computational, deterministic structures an agent queries through a context-as-a-service layer, a third source of knowledge alongside training data and the vector database. Served this way, guide gives an agent the cross-domain knowledge it needs to get the answer right the first time. Sensor Sensor,…
Shopify powers observability for global-scale commerce with ClickHouse (opens on the source site)
Shopify unified global-scale observability on ClickHouse, achieving up to 30x faster queries while ingesting 100 million events per second at peak.
The search multiplier: Driving revenue, productivity, and AI at scale (opens on the source site)
Elastic ·
An independent IDC study found organizations deploying the Elasticsearch Platform for enterprise search and agentic AI deliver AI products faster, drive higher revenue, boost productivity, and strengthen operational resilience across the business.
How Sony LIV uses ClickHouse Cloud to deliver live streaming analytics at billion-row scale (opens on the source site)
Sony LIV consolidated fragmented batch, Elasticsearch, and BigQuery workloads on ClickHouse Cloud, delivering sub-second analytics across billions of daily streaming events.
New in Confluent Cloud and WarpStream: Evolving the Data Streaming Platform for AI, Scale, and Control (opens on the source site)
Accelerate enterprise streaming and AI workloads with new product updates across Kora, connectors, Flink, Tableflow, AI anomaly detection and forecasting, security enhancements, and Warpstream.
Creativity at Enterprise-Scale, Without Compromise: The OFF+BRAND. Story (opens on the source site)
Codrops ·
Telling stories on some of the biggest canvases on the web, from Awwwards Site of the Year to the front door of Microsoft.
Govern Replit at scale (opens on the source site)
Replit ·
New Admin API, Audit Logs, and Workspace Settings give organizations more insight and flexibility into how they adopt AI with Replit. Replit helps teams turn ideas into working software quickly. Most companies start with a handful of builders and a few prototypes. Then it works, and it spreads: more teams, more projects, more software running in production. That shift creates work for a specific group of people. As AI tools spread across a company: IT, procurement, and admin teams absorb the cost. They field permission requests, run access reviews, make policy decisions tool by tool, and…
A study of sequence weighting at scale (opens on the source site)
TL;DR: We study the scaling laws of data weighting across in-house and open-weight LMs, finding non-monotonic behavior across scales. We vary the weight assigned to sequences during training and measure how strongly the model’s loss reduction on a sequence depends on the sequence’s weight. Taken together, our results are consistent with a general trend: as models transition from small to medium scale, they transition from learning general patterns independent of data weight to learning data-specific patterns proportional to the data weights. As models then transition from medium to large…
Amazon DynamoDB now supports real-time vector search at any scale (opens on the source site)
AWS ·
DynamoDB now supports native vector search with single-digit millisecond latency at 99%+ recall. It is designed for any scale, even trillions of vectors and requires zero infrastructure management.
GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model (opens on the source site)
Facebook ·
Meta’s Generative Ads Recommendation Model (GEM), the foundation model behind ads recommendations across Instagram and Facebook, now trains at LLM scale on several thousand of the latest-generation GPUs. This post goes into the details on how we achieved: doubling end-to-end (E2E) training efficiency to 20–25% Model FLOPs Utilization (MFU) while scaling training FLOPs 4x in [...] Read More... The post GEM Training: How Meta Doubled the Efficiency of Its LLM-Scale Ads Foundation Model appeared first on Engineering at Meta.
Eval-driven development: Lessons from evaluating GenAI at scale (opens on the source site)
Airbnb ·
How Airbnb teams build trustworthy Generative AI products by treating evaluation as a first-class engineering discipline; not an afterthought.Nestled into the lush hillside, this stunning modern retreat features striking natural wood architecture, terraced balconies, and a serene landscape.By: Rohit Girme, Dan Miller, Mia Zhao, Lifan Yang, Clint KellyIntroductionGenerative AI breaks a lot of the assumptions that used to hold true for software testing. Unlike traditional software, LLM outputs are non-deterministic, and “correct” is subjective. Because so much judgment is involved, you often…
Agent platform (Part 1): How we help Grab build and run AI agents at scale (opens on the source site)
Grab ·
Part 1: From one support bot to a framework At Grab, AI agents have evolved from interesting team prototypes into production services used every day by millions of merchants, drivers, and consumers. Today, more than 500 services run on our internal agent framework, over 50 Model Context Protocol (MCP) servers are registered on our remote MCP framework, and a single Large Language Model (LLM) gateway fronts every model call across the company, handling billions of tokens each month. None of this was designed up front. It began as the plumbing behind one internal support bot, which then…
Securing Infrastructure at Scale: Introducing Pinterest’s Resource Provisioner Pipeline (RPP) (opens on the source site)
Ammar Ekbote | Senior Software EngineerChan Kim | Senior Software EngineerManaging Infrastructure as Code (IaC) across a massive organization comes with a unique set of security and logistical challenges, particularly when operating within a distributed, multi-repository architecture. At Pinterest, we designed the Resource Provisioner Pipeline (RPP), our specialized, proprietary Terraform execution engine to safely manage both critical and non-critical infrastructure changes.In this post, we will look under the hood of the first iteration of the RPP system. We will explore how it established…
Related topics
This page is generated automatically from the engineering blogs we follow. Every post links to its source, where it was published. See all sources.