Log inSign up
LMCache Lab
266 posts
@lmcache

LMCache Lab

@lmcache
🧪 Open-Source Team that maintains LMCache and Production Stack 🤖 Democratizing AI by providing efficient LLM serving for ALL
Github, Online
lmcache.ai
Joined September 2024
54
Following
1,988
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @lmcache
    LMCache Lab
    @lmcache
    2h
    🔬Calling ML engineers: turn your KV-cache editing idea into a real benchmark with LMCache! KV-cache techniques such as token dropping, KV cache editing, or even cartridge-style KV cache training allows an LLM to decode much faster or produce much better text outputs. But how to
    A whiteprint-style engineering drawing in violet-blue line on bone paper, laid out in two horizontal lanes. The upper lane, labeled INFERENCE ENGINE — GPU, is a long open housing holding two blocks filled with vermilion halftone: one at the left, one at the right. Between them the lane is empty paper. A dashed rule runs the full width below the housing, marking the device boundary. The lower lane, labeled LMCACHE —...
    From Research Paper to Real Workload: KV-Cache Editing with the LMCache SDK
    From blog.lmcache.ai
  • @lmcache
    LMCache Lab
    @lmcache
    9h
    1/ Tomorrow's LMCache Community Office Hour is an AMA with Karsten Wade (@quaid), our Open Source Community Architect. For two months he's been doing a read-only assessment of the community: observing, not interrupting. Wed Sep 9, 11am PT. lmcache-officehours.zapier.app
    1
  • @lmcache
    LMCache Lab
    @lmcache
    23h
    Tomorrow's LMCache Community Office Hour is an AMA. Karsten Wade (@quaid) has spent two months quietly reading the project — PRs, issues, Slack threads, plus an agentic scan. He'll share what he found, then answer anything you ask. Wed Sep 9, 11am PT lmcache-officehours.zapier.app
  • @lmcache
    LMCache Lab
    @lmcache
    Aug 21
    𝗔𝗴𝗲𝗻𝘁𝗶𝗰 𝗶𝗻𝗳𝗲𝗿𝗲𝗻𝗰𝗲 𝗶𝘀 𝘁𝘂𝗿𝗻𝗶𝗻𝗴 𝗞𝗩-𝗰𝗮𝗰𝗵𝗲 𝗺𝗮𝗻𝗮𝗴𝗲𝗺𝗲𝗻𝘁 𝗶𝗻𝘁𝗼 𝗮 𝗳𝗶𝗿𝘀𝘁-𝗰𝗹𝗮𝘀𝘀 𝘀𝘆𝘀𝘁𝗲𝗺𝘀 𝗽𝗿𝗼𝗯𝗹𝗲𝗺. @SemiAnalysis_ just released 𝗔𝗴𝗲𝗻𝘁𝗫, a benchmark built around real agentic coding workload patterns: long multi-turn
    Image
    2
  • @lmcache
    LMCache Lab
    @lmcache
    Aug 21
    🚀 𝗟𝗠𝗖𝗮𝗰𝗵𝗲 𝘃𝟬.𝟱.𝟰 𝗶𝘀 𝗼𝘂𝘁! 🗂️ 𝗖𝗼𝗼𝗿𝗱𝗶𝗻𝗮𝘁𝗼𝗿 𝗮𝘀 𝗮 𝗳𝗹𝗲𝗲𝘁 𝗰𝗼𝗻𝘁𝗿𝗼𝗹 𝗽𝗹𝗮𝗻𝗲 (𝗠𝗣 𝗺𝗼𝗱𝗲) • Cache events flow through a dedicated ingest layer into fleet controllers, with unified per-tier L1 usage tracking, metrics export, and fleet-wide
    Image
    Made with AI
Advertisement
Advertisement