Search

Tag: #efficiency28 results

LightMem Slashes LLM Memory Costs

LightMem Slashes LLM Memory Costs

LightMem is a lightweight memory system for LLMs that reduces token usage, API calls, and latency without sacrificing accuracy in long conversations and multi-task agents. It addresses redundancy in dialogues, rigid segmentation, and costly online updates by mimicking human hierarchical memory mechanisms. The paper is accepted to ICLR 2026 with open-source code available.

机器之心MediaFeb 26#memory-augmented#llm-agents#efficiency
MiniMax: Bubble or AI Future?

MiniMax: Bubble or AI Future?

MiniMax shines in multi-modal models with top cost-performance, efficient operations, and rapid iterations. It prioritizes to-C entertainment products, global markets, and light-asset APIs, avoiding chatbot wars. Competitive pricing positions it well for Agent scaling.

Page 2 of 3