Memory
We track 43 posts about Memory from 29 engineering blogs. Most active: Android, Twilio, Learnk8s. Latest post: Oct 9, 2026.
Companies writing about Memory
Recent posts
Vue 3 Composition API: Stop Watcher Memory Leaks in Composables (opens on the source site)
How to manage watcher lifecycles and cleanup in Vue 3 composables, preventing memory leaks when watchers are created outside synchronous setup. Continue reading Vue 3 Composition API: Stop Watcher Memory Leaks in Composables on SitePoint.
Part 4: Safety and governance for LLM systems: guardrails, PII, audit, and memory (opens on the source site)
The level where an LLM system stops being a demo and earns the right to touch real data and real decisions: layered guardrails that fail closed, PII handled at the boundary, an immutable audit trail, and scoped memory. By the time an LLM system is making decisions that matter, “it usually works” is no longer the bar. This is Level 4 of the maturity model — safety and governance — and it’s where four disciplines that teams tend to bolt on late have to be designed in instead. They share one idea: don’t trust a single point to do the right thing. Layer independent guardrails so a miss at one is…
Go CPU and Memory Requests and Limits in Kubernetes (opens on the source site)
Learnk8s ·
Set Kubernetes CPU and memory requests and limits for Go services by checking GOMAXPROCS, GOMEMLIMIT, live heap, goroutine stacks, CPU throttling, and garbage collection cost.
Memory safety for Postgres extensions in C/C++ (opens on the source site)
How ClickHouse’s Postgres extensions handle the memory safety challenges of combining C and C++, from clean language boundaries to isolated helper processes.
Trading off compute for memory with activation checkpointing (opens on the source site)
The following is part of a series of posts about 2026 summer intern projects—for more, see “What the interns have wrought, special jumbo 2026 edition”
chDB Durable Layer for agent memory (opens on the source site)
The chDB Durable Layer keeps agent memory fast to query locally while making its analytical state recoverable across laptops, CI jobs, and short-lived sandboxes.
Building Multi-Tier AI Agent Memory with TypeScript and SQLite-vec (opens on the source site)
Architect a tiered local memory system for AI agents using TypeScript, better-sqlite3, and sqlite-vec across episodic, semantic, and procedural stores. Continue reading Building Multi-Tier AI Agent Memory with TypeScript and SQLite-vec on SitePoint.
Vercel Sandbox now supports memory observability (opens on the source site)
Vercel ·
Vercel Sandbox observability now includes memory usage data. You can access sandbox memory usage data in the dashboard and through the CLI via the vercel metrics command. Sandbox observability memory in the dashboard The Memory Usage card reports average, P75, and P95 memory across your sandboxes, alongside the existing CPU usage and data transfer metrics on the sandbox overview page and the project and team-level pages. On the sandbox detail page, charts are designed to show how close a sandbox is running to its memory ceiling by: Automatically scaling the y-axis to sandbox's memory limit…
Creating a memory dump in C# (opens on the source site)
.NET ·
Learn how to create a memory dump in C# to capture the state of your application for debugging purposes. The post Creating a memory dump in C# appeared first on .NET Blog.
Cutting Node.js Memory Footprint with HyperLogLog and Count-Min Sketch in TypeScript (opens on the source site)
Ingest and aggregate millions of streaming events in Node.js with sub-megabyte RAM using HyperLogLog and Count-Min Sketch written in pure TypeScript. Continue reading Cutting Node.js Memory Footprint with HyperLogLog and Count-Min Sketch in TypeScript on SitePoint.
Size-Specialized Memory Allocation (opens on the source site)
Go ·
The Go Blog Size-Specialized Memory Allocation Michael Matloob 16 September 2026 Go 1.27 includes faster memory allocation for allocations of 80 bytes or fewer. Allocations can be up to 20-30% faster, making allocation-heavy programs up to 1% faster. The Go runtime improves the performance of those allocations by adding specialized functions that are used to allocate certain sizes. These specialized functions can then make certain assumptions that make them faster and easier to optimize. This blog post will explain how this works and how it makes your programs faster. Heap allocations are…
How to Orchestrate Multi-Call Conversations with an LLM and Twilio Conversation in Node.js Memory (opens on the source site)
Twilio ·
Learn how to orchestrate multi-call voice conversations using Node.js, OpenAI, and Twilio Conversation Memory to persist caller context across separate calls.
How to Orchestrate Multi-Call Conversations with an LLM and Twilio Conversation Memory in Python (opens on the source site)
Twilio ·
Learn how to build a Python FastAPI voice agent that uses Twilio Conversation Memory to remember callers across separate phone calls, so if someone hangs up and calls back, the agent picks up right where the conversation left off.
How to Orchestrate Multi-Call Conversations with an LLM and Twilio Conversation Memory with PHP (opens on the source site)
Twilio ·
In this tutorial you'll make a PHP service using Open Swoole that retains caller context, preferences and action history across multiple separate inbound calls.
Persistent memory for eve agents (opens on the source site)
Vercel ·
eve agents can now retain context across sessions and use it in future conversations. Persistent memory is organized into slots. You can define named slots in files under agent/memory/. Each slot specifies a provider, which stores and retrieves the memory, and a scope, which determines who or what shares it. For example, you can keep separate memory for each authenticated user. Before each turn, eve retrieves relevant memory and adds it to the model's context. Depending on the provider, memory can be updated automatically after a turn, through tools the agent uses, or through both methods. To…
Java JVM CPU and Memory Requests and Limits in Kubernetes (opens on the source site)
Learnk8s ·
Setting Kubernetes CPU and memory requests and limits for a JVM service is four coupled decisions: container memory, heap size, GC selection, and CPU quota. Use the calculator to explore the space.
Elevating app quality: Reducing memory usage and improving device migration (opens on the source site)
Android ·
Posted by Raghavendra Hareesh Pottamsetty, GM, Google Play Developer & MonetizationMaintaining a healthy Android ecosystem is a shared commitment where every app and game has a role to play. To help you deliver the premium experiences users expect, Google Play is introducing two new quality requirements: one focused on reducing app memory footprint, and another on providing a secure, seamless device migration experience. First, to help developers navigate industry-wide hardware constraints and Android's broader memory limits, Google Play is establishing new performance thresholds. Second, as…
How we saved 100 terabytes of memory by optimizing 1.1.1.1’s DNS cache (opens on the source site)
Five Rust-level memory optimizations to the DNS cache layout of Big Pineapple cut per-entry memory by 56%, freeing approximately 100 TB of memory across Cloudflare's fleet.
NVIDIA NVLink Fusion Expands With NVHBM Custom High-Bandwidth Memory (opens on the source site)
Nvidia ·
The next wave of AI is placing new demands on infrastructure. As AI agents and trillion-parameter workloads become mainstream, the performance of AI infrastructure depends not only on compute, but on how compute, memory, storage, networking and software are designed together as a unified system. To help hyperscalers and AI innovators build the next generation […]
Elevating app quality: Reducing memory usage and improving device migration (opens on the source site)
Android ·
Posted by Raghavendra Hareesh Pottamsetty, GM, Google Play Developer & Monetization Maintaining a healthy Android ecosystem is a shared commitment where every app and game has a role to play. To help you deliver the premium experiences users expect, Google Play is introducing two new quality requirements: one focused on reducing app memory footprint, and another on providing a secure, seamless device migration experience. First, to help developers navigate industry-wide hardware constraints and Android's broader memory limits, Google Play is establishing new performance thresholds. Second, as…
Inside LinkedIn's cognitive memory agent for agentic personalization (opens on the source site)
Ryan is joined by Praveen Bodigutla, Principal AI Researcher at LinkedIn, to chat about the four-layer memory system his team built to give LinkedIn's hiring assistant a persistent, personalized state.
Preparing your app for broader memory limits (opens on the source site)
Android ·
Posted by Blair Harmon, Director of Product Management, Android PlatformA great user experience is central to Android's mission, and delivering on that promise requires keeping devices fast, responsive, and reliable. This is why memory optimization is more critical than ever. Across the ecosystem, new devices are maintaining or even decreasing their physical memory capacity in response to memory price increases, yet users continue to expect the same seamless, high-performance app experience. In Android 17, we introduced per-app memory limits, starting with Pixel devices, to help protect the…
Preparing your app for broader memory limits (opens on the source site)
Android ·
Posted by Blair Harmon, Director of Product Management, Android Platform A great user experience is central to Android's mission, and delivering on that promise requires keeping devices fast, responsive, and reliable. This is why memory optimization is more critical than ever. Across the ecosystem, new devices are maintaining or even decreasing their physical memory capacity in response to memory price increases, yet users continue to expect the same seamless, high-performance app experience. In Android 17, we introduced per-app memory limits, starting with Pixel devices, to help protect the…
How do functions like alloca allocate memory from the stack? (opens on the source site)
The usual probing. The post How do functions like alloca allocate memory from the stack? appeared first on The Old New Thing.
Pandas 3 string dtype: PyArrow strings cut memory use (opens on the source site)
Using Python Pandas 3? Strings use PyArrow, not Pandas 2’s Python strings (dtype “object”): The post Pandas 3 string dtype: PyArrow strings cut memory use appeared first on LernerPython.
Build Persistent Customer Memory with Twilio Agent Connect and Conversation Intelligence (opens on the source site)
Twilio ·
Learn how to build an AI assistant with persistent customer memory using Twilio Agent Connect, Flex, and Conversational Intelligence.
Python list memory: Why sys.getsizeof ignores the elements (opens on the source site)
How much memory does a Python list use? sys.getsizeof will tell you, sort of: It reports the memory used by the list, but not its elements: The post Python list memory: Why sys.getsizeof ignores the elements appeared first on LernerPython.
Don’t stop early: Case-folding source code at memory speed (opens on the source site)
GitHub ·
How a branch-free loop and byte-space arithmetic let GitHub case-fold every byte of code search at >45 GiB/s on a single core. The post Don’t stop early: Case-folding source code at memory speed appeared first on The GitHub Blog.
Knowledge as Code: The Memory File Just Got a Spec (opens on the source site)
Pulumi ·
Five weeks ago I wrote that the least glamorous piece of an agent loop is also the one that decides whether it compounds: memory. A markdown file outside the context window that holds what is done, what is next, and what was learned, because the model forgets all of it between runs. Write the memory file before the loop. What I left open, because there was nothing to point at, was the format. My memory file looked nothing like yours, and neither of our agents could read the other’s. Three days after that post went live, Google shipped an answer. The pattern everyone copied Andrej Karpathy…
Testing Java Memory Management with Chronicle-FIX using AI (opens on the source site)
While I am sceptical of using AI for release code, it has plenty of uses that previously weren’t practical, such as determining how easy your software is to use. If an AI can “figure it out” with a few hints, then you are on the right track. For me, the value of AI is what you learn using it. For more Techincal Information on Chronicle-FIX What AI Does Well and What It Doesn’t Claude and Codex are effective for producing idiomatic code; for low-latency code, it needs a significant body of example code. In this case, it was able to utilise sample code for benchmarks. If it was being used to…
Related topics
This page is generated automatically from the engineering blogs we follow. Every post links to its source, where it was published. See all sources.