Log in
Sign up
Products
Tensormesh Inference
Serverless and Reserved inference with built-in context caching.
Tensormesh Operator
Bring production context caching to your on-prem Al stack.
Coming Soon
Products
Tensormesh Inference
Tensormesh Operator
Pricing
Partners
Company
About us
Learn how we’re ending GPU waste and the "Amnesia Tax"
Events
Connect with our team at upcoming summits.
Careers
Help us build the persistent memory layer for AI.
Company
About Us
Events
Careers
Resources
Blog & News
Insights on reducing GPU costs and improving latency.
Documentation
Technical guides for OpenAI-compatible integration.
LMCache
The open-source engine powering our core technology.
Savings Calculator
Quantify when caching beats recomputing and by how much.
FAQ
Find answers to common questions about the platform.
Resources
Blog
Documentation
LMCache
Savings Calculator
FAQ
Contact
Talk to Sales
Connect with an engineer to solve your GPU bottleneck.
Partner Inquiry
Explore opportunities to partner and grow together
Become a partner
Contact sales
Log in
Sign up
Blog
Insights & Updates
All
News
Articles
Shown:
0
Latest News
July 23, 2026
Tensormesh and AMD Collaborate to Empower Fewer GPUs to Serve More Models
Junchen Jiang
CEO, Co-Founder
Read Now
News
All
May 27, 2026
Tensormesh Raises $20M from Investors Including AMD Ventures, CoreWeave, NVentures, Launches Tensormesh Inference to Fix AI’s Most Expensive Problem
Junchen Jiang
CEO, Co-Founder
Read Now
News
All
October 23, 2025
Tensormesh Emerges From Stealth to Slash AI Inference Costs and Latency by up to 10x
Junchen Jiang
CEO, Co-Founder
Read Now
News
All
Articles
May 13, 2026
The AI Agent Metrics That Actually Matter: Beyond Tokens and Latency
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
May 6, 2026
Tensormesh Inference: Cheaper LLM Inference for AI Agents
Sandro Mazziotta
Head of Product Management
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
April 29, 2026
Agentic AI Inference Cost: How LLM Agent Loops Break Caching and Drain Your Budget
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
April 28, 2026
Inside Tensormesh: Meet our CTO and Chief Scientist
Kuntai Du
Chief Scientist, Co-Founder
Yihua Cheng
CTO, Co-Founder
Read article
Articles
All
April 22, 2026
Enterprise AI Vendor Lock-In: What It Costs When Your Provider Pulls Access
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
April 15, 2026
Introducing Tensormesh Beta 2.2: Serverless Inference & $0 Cached Input Tokens
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
April 8, 2026
How We Optimized Redis for LLM KV Cache: 0.3 GB/s to 10 GB/s
Samuel Shen
Software Engineer
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
February 25, 2026
Introducing Tensormesh Beta 2: One-Click LLM Deployment, New UI & Real-Time Cost Savings
Bryan Bamford
Marketing, Enterprise and Partnerships
Read article
Articles
All
February 18, 2026
Agent Skills Caching with CacheBlend: Achieving 85% Cache Hit Rates for LLM Agents
Kuntai Du
Chief Scientist, Co-Founder
Read article
Articles
All
Show more