-
Consistency compounds: Valkey's journey to 200 Gbps
Caching, Valkey -
The concurrency cliff is a memory limit
Inference -
Your KV cache benchmark is “hi hi hi”
Inference -
vLLM's Hash Chain and Why Prefix Caching Is Still Prefix Caching
Inference -
Disaggregation makes KV cache a system primitive
Inference -
KV Caching Pays Off Under Load
Inference -
Beyond the Goals, Three Ways Momento Scales the Football World Cup in Real Time
Media & Entertainment -
A New Live Streaming Origin Built for Global Scale
Media & Entertainment -
Introducing valkey-lab: Stop Guessing When Your Cache Hits Its Limit
Valkey -
Why Snap Was Willing to Fork, and Why They Still Came Back
Caching, Valkey -
Why Large Payloads Break Caches at Scale
Scale -
Disaggregated LLM Inference, Part 3: Why Your Networking Stack May Not Be Ready
AI/ML -
Disaggregated Inference,Part 2: Moving the KV Cache Without Stalling the Decode
AI/ML -
The Snowflake Moment for Inference

AI/ML -
Disaggregated Inference, Part 1: When & Where to Route
AI/ML -
Prefill and Decode Want Different Chips. The Economics Finally Agree.
AI/ML, Performance -
1-Bit Models Just Moved the Pareto Frontier

AI/ML -
Your AI Remembers Everything Except the Thing You Keep Telling It
AI/ML -
KV Cache Isn't a Caching Problem
AI/ML -
The Rise of the Internal Cache Platform
Caching -
A Roadmap for KV Cache Offloading at Scale
AI/ML -
GPUs are the most expensive resource in tech. We’re using them badly.
AI/ML -
Stop CDN Leeching with Concurrency Tracking
Media & Entertainment, Security -
What Hyperscale Caching Taught Us About GPU Utilization
AI/ML -
Tooling is a Scaling Strategy
Scale -
Understanding the NxM Problem in Distributed Caches
Caching -
Why Large Cache Systems Need Routing Layers
Caching -
Reduce TTFT by >50% with LMCache + Momento


AI/ML -
Reduce TTFT by >50% with LMCache + Momento Accelerator

Caching, Performance -
Performance Engineering Lessons from the Unlocked Conference

Events -
Large Objects Ruin the Party - Valkey 9 Tames Them

Caching, Valkey -
The Real Cost of Swapping Infrastructure
Infrastructure -
Breakthroughs Are Just Boring Improvements That Pile Up
Caching, Valkey -
Cache Rebalancing Was Broken. Here's How Valkey 9.0 Fixed It
Caching, Valkey -
The Momento Platform


Product Update -
Designing smarter caches with Valkey 9.0’s numbered databases
Valkey -
Cache It - Episode #7 - Valkey 9.0: Databases, Clustering, and Details with Kyle Davis
Podcast -
Valkey 9.0 - The Next Generation of Caching

Valkey -
The 5 Metrics that Predict Cache Outages
Caching -
The Latest Redis Vulnerability Exposes a Bigger Problem

Infrastructure -
Valkey 8.1 vs Redis 8.2: Memory Efficiency at Hyperscale
Valkey -
Momento is the DNA of AI agents
AI/ML -
(Buffer-Free Video)^AI 2025: Where AI and Video Innovation Came to Life
Events -
Valkey Turns One: How the Community Fork Left Redis in the Dust
Perspectives -
FOX Monitors Super Bowl Viewership Experience with Real-time Data Insights

Case Study -
NAB 2025: Must-see Sessions
Events -
Momento Leaderboards Just Got Even Better with Competition Ranking
Product Update -
RaiderlO effortlessly powers leaderboards for millions of World of Warcraft players with Momento

Gaming -
The Future of Video Streaming Demands More Than Just Compression
Media Storage -
Elevate Your Live Streaming: A Smooth Transition from AWS Elemental MediaStore to Momento Media Storage
Media Storage -
Momento Cache – Caching you can trust for data reliability
Caching -
Momento: A platform for everyone
Perspectives -
S3: The greatest AWS service ever, is not a Live Media Origin
Perspectives -
How to build a real-time notification system with Momento: a step-by-step guide
Topics -
How to build a leaderboard service with Momento: a step-by-step guide
Leaderboards -
How to build a real-time chat application with Momento: a step-by-step guide
Topics -
What is real-time data processing?
Performance -
Integrating Amazon DynamoDB Streams with Momento, via Amazon EventBridge: Fully Automated via AWS CDK!
Integration -
Building a real-time weather update system with Momento, Amazon EventBridge, and DynamoDB
Integration -
Announcing Momento's Series A and the General Availability of Topics

Launch -
Reflecting on Momento's Inclusion in the InfraRed100


Perspectives -
Is The Serverless Fairytale Over?
Serverless -
Momento’s security pillars for secure apps
Security -
It’s official - we’re verified on Postman!
Security -
Three key questions to pave the way for success with serverless
Serverless -
How to fix connection timeout issues with AWS Lambda in VPCs
Engineering -
Serverless at enterprise scale: What it is and where to begin
Serverless -
Seamlessly caching MongoDB Atlas (it's automagic!)
Database -
Momento isn’t serverless Redis. It’s caching reimagined for a post-Redis world.
Caching -
RIP Redis: How Garantia Data pulled off the biggest heist in open source history

Perspectives -
Lessons learned at GDC: Our thoughts on the next generation of game dev
Gaming -
When to use sorted sets vs Momento Leaderboards
Leaderboards -
Yes, we built a multiplayer, squirrel-themed replica of Flappy Bird on Momento
Gaming -
How to develop a chat app with built-in moderation
Event-Driven Architecture -
Introducing RoboMo—A Discord bot powered by Momento Vector Index
AI/ML -
A tale of gRPC keepalives in the Lambda execution context
Caching -
Observe It - Episode #5 - How high cardinality impacts usability and cost with Alex Kehlenbeck
Podcast -
Observe It - Episode #4 - The magic of API observability with Jean Yang
Podcast -
Linear Scaling and the End of the Rainbow
Strategy -
Scaling Strategies for Cloud Applications
Strategy -
Serverless Applications 101: A Basic Overview and Best Practices
Serverless -
ICYMI - Leafing through Momento’s fall updates
Product Update -
4 key ingredients to building for scale
Perspectives -
ElastiCache Serverless has a hidden feature: Memcached replication

Caching -
Is S3 Express One Zone a serverless cache?

Caching -
WebSockets Guide: How They Work, Benefits, and Use Cases
Topics -
Observe It - Episode #3 - Reducing observability costs with Kevin Lin
Podcast -
Distributed locks: Save 50%+ on your DynamoDB lock client
Integration -
Highlights from the launch of Amazon ElastiCache Serverless
Caching -
Distributed locks: Making them easier with Momento
Integration -
How we turned up the heat on Node.js Lambda cold starts
Performance -
DynamoDB Magic Numbers: Optimize your DynamoDB spend with simple math

Database -
Observe It - Episode #2 - OpenTelemetry: yay or nay?
Podcast -
Did you say you want a distributed rate limiter?
Rate-Limiting -
Momento just got more powerful: Introducing Topics

Topics -
Momento Topics just got more secure: introducing embedded token identifiers

Topics -
Supercharge your default Drupal cache by integrating with Momento
Integration -
Best practices for ElastiCache Redis autoscaling–and how to do better
Caching -
Unity chat demo: Quickly build a multiplayer chat with serverless pub/sub
Gaming -
Introducing Momento Leaderboards: the serverless leaderboard service
Gaming -
Momento Cache is the cloud-native answer to ElastiCache Redis
Caching -
Single table design for DynamoDB: The reality
Database -
Horizontal scaling with ElastiCache Redis: Stop getting burned by hot keys and shards
Caching -
3 crucial caching choices: Where, when, and how
Caching -
Quick Primer on ElastiCache Redis Maintenance Windows
Performance -
Cache-it - Episode #6 - Agent Centric LLMs with Langroid starring Prasad Chalasani
Podcast -
Optimizing Sequelize: Using a built-in read-aside cache for peak efficiency
Caching -
Observe It - Episode #1 - A deep dive into Observability with Ben Sigelman
Podcast -
Exploring leaderboards with competition ranking
Caching -
6 common caching design patterns to execute your caching strategy
Caching -
Cache-it - Episode #5 - Simplifying event-driven architectures with Eric Johnson
Podcast -
Momento: A front-end developer’s best kept secret
Front-End Developement -
ICYMI - Momento’s hot squirrel summer
Product Update -
API keys vs tokens - what’s the difference?
Security -
Cache-it - Episode #4 - Million dollar lines of code: Engineering your cloud cost optimization with Erik Peterson
Podcast -
Redis compatibility clients: Your pathway to Momento

Integration -
Building an interactive live reaction app with Next.js and Momento🎯
Topics -
Introducing the Momento Token Vending Machine
Security -
What is a vector index?
Database -
How we built Momento Topics, a serverless messaging service
Topics -
Momento feature discussion: Fine-grained access control and HTTP support with Daniela Miao and Allen Helton

Product Update -
Why are WebSockets so hard?
Topics -
MoCon 2023: A Day of Inspiration and Innovation
Events -
Cut the caching clutter: understanding cache types
Caching -
Chatting on the Edge: Integrating Momento with Netlify and Vercel
Integration -
Cache-it - Episode #3 - Vector databases, single-table design, and beyond with Alex DeBrie
Podcast -
Adding chat functionality to your games and apps

Topics -
Cache-it - Episode #2 - Indexing adventures in the age of embeddings: Building a world-class search system
Podcast -
Cache-it - Episode #1 - Applying lessons from caching to ML feature stores with Yao Yue
Podcast -
Why tail latencies matter
Caching -
Momento Cache is now accessible at the edge with Cloudflare
Caching -
Turbocharging Pelikan Cache on Google Cloud’s latest Arm-based T2A VMs

Performance -
I built a 3.75-million subscriber chat system in an afternoon

Topics -
Momento is now fully integrated into the LangChain Ecosystem
Integration -
Build on Momento: IoT device status
Caching -
Hello World! Introducing the Momento Web SDK
Launch -
Now available: Momento Bulk Writer
Product Update -
Build on Momento: Instant messaging
Caching -
Easy mode: Drop Momento right into your Redis app
Integration -
Announcing AWS PrivateLink connectivity for Momento
Launch -
Momento Cache vs. Redis: the key differences
Caching -
Momento Console is here

Launch -
How caching fits into your Amazon Aurora scaling strategy
Database -
Build on Momento: Event routing with Momento Topics
Event-Driven Architecture -
Real World Serverless Podcast: Kirk Kirkconnell
Podcast -
Build on Momento: How we made instant messaging for Acorn Hunt
Caching -
Improve app performance by caching at every layer

Caching -
Major release: v1.0 of the Momento Go client
Product Update -
Momento flush cache is a modern convenience
Caching -
Build on Momento: Tips and tricks for a lightning-fast game leaderboard
Caching -
Distributed SQL, distributed cache
Database -
Major release: v1.0 of the Momento Python client
Product Update -
Boost your stats with our new increment API
Product Update -
Big news, India: We’re here!
Launch -
What Is Serverless Podcast
Podcast -
Major release: v1.0 release of the Momento Node.js Client!
Product Update -
Did someone say collection data types? Oh, we did!
Product Update -
Customer-led growth: Adding collection data types to Momento

Caching -
We're back with another collection data type: Sorted sets!
Product Update -
Maximize cost savings and scalability with an optimized DynamoDB secondary index
Database -
Software Daily Podcast
Podcast -
Which flavor of DynamoDB secondary index should you pick?
Database -
You’re asking “why serverless?” Let’s talk!

Serverless -
Database caching: Improved hit rates courtesy of DynamoDB Streams
Caching -
Database caching: Outdoing DAX with the serverless solution DynamoDB deserves

Caching -
Database caching: Momento Cache is the easiest way to cache DynamoDB
Caching -
Better together: serverless and multi-cloud
Serverless -
Fascinating facts about facades at CBS Sports
Case Study -
What really matters in DynamoDB data modeling?
Database -
Reducing cloud infrastructure cost is easier than you think
Caching -
Fighting off faux-serverless bandits with the true definition of serverless

Serverless -
Major release: v1.0 of the Momento .NET client
Product Update -
Shockingly simple: Tuning Momento’s .NET cache client
Caching -
Effortless caching reducing MongoDB operations
Database -
Moving your bugs forward in time
Perspectives -
Using a cache to accelerate DynamoDB (or replace it)
Caching -
We built a serverless cache you can add to your stack before lunch

Perspectives -
How a centralized cache elevates serverless applications
Caching -
A spooky tale of overprovisioning with Amazon DynamoDB and Redis
Caching -
Oops, Momento ate 98% of my GCP Cloud Run and Firestore latencies!
Performance -
Shockingly simple: Tuning Momento’s Python cache client
Caching -
Exceptions are bugs
Perspectives -
Shockingly simple: Tuning the Momento JavaScript cache client
Performance -
Faster APIs, faster developers: API Gateway custom authorizers
Performance -
Open Source Startup Podcast
Podcast -
Simple. The way cloud pricing should be.
Serverless -
Think before you cache
Caching -
OpenTelemetry: Tips to navigate the sea of observability options
Perspectives -
The dark art of multi-tenancy
Perspectives -
Shockingly simple: Cache clients that do the hard work for you
Caching -
Oops, Momento ate 60% of my Lambda latencies!
Performance -
Stop settling for almost serverless
Serverless -
The biggest miss in your cache hit rate

Caching -
Finally, a serverless cache that delivers on the promise of the cloud era
Caching -
Making Pelikan fly on Arm: Diving deeper into our adventures with Tau T2A VMs

Performance -
4 tips for building high-performance systems

Performance -
Real World Serverless Podcast: Khawaja Shams
Podcast -
Serverless Cache: The missing piece to go fully serverless
Caching -
Engineering Founders Podcast
Podcast -
Memcached vs. Redis: When scalability and reliability matter
Caching