GPU efficiency

Subconscious Raises $5.1M for AI Agent Inference Efficiency

Subconscious Raises $5.1M for AI Agent Inference Efficiency

Subconscious has raised $5.1 million to commercialize an MIT-derived inference platform that dramatically reduces the cost and latency of long-running AI agents. The platform uses dynamic context compression to cut token consumption by up to 82% in production, with no model or application changes required. For engineering teams hitting AI spending ceilings and regulated-sector buyers needing air-gapped deployments, the timing is pointed.

Subconscious Raises $5.1M for AI Agent Inference Efficiency Read More »

MinIO MemKV: Purpose-Built AI Inference Cache Storage

MinIO MemKV: Purpose-Built AI Inference Cache Storage

MinIO has launched MemKV, a purpose-built KV cache storage product targeting NVIDIA’s G3.5 memory tier via NVMe/RDMA. The product promises a 75x improvement in inference time-to-first-token and up to $2M in annual GPU efficiency savings for a typical enterprise deployment. ECI Research examines the business case, technical architecture, and what buyers need to validate before committing.

MinIO MemKV: Purpose-Built AI Inference Cache Storage Read More »