Dileeshvar Radhakrishnan

MinIO Blog Posts

Prompt Caching: Stop Paying GPUs to Read the Same Prompt Twice
arrow
Every coding assistant, enterprise chatbot, and agent loop resends the same tool schemas, system instructions, and policy text on every turn, and the GPU rebuilds all of it into attention state before it can emit a single output token. Prompt caching computes that prefix once and reuses the KV state, and MemKV takes it from a per-process optimization to a shared NVMe-backed tier that survives routing across replicas, HBM eviction, worker restarts, and concurrency.
Agent Memory
AI/ML
AIStor
Performance
Operations
From Cache Hits to Production SLAs | Part 3 of 3
arrow
Cache hits are not the outcome. Part 3 turns the architecture from Parts 1 and 2 into an evaluation framework: capture an honest tail-latency baseline first, treat MinIO's published 53-second-to-703-millisecond TTFT result as a proof point to reproduce rather than a business-case input, and follow the measurement chain from repeated context through prefix reuse and avoided recompute to unit economics, ending in a buyer checklist that judges a shared context tier on P99 TTFT and jitter rather than throughput or capacity.
Agent Memory
AI/ML
AIStor
Operations
Performance
When Repeated Context Becomes an Infrastructure Problem | Part 2 of 3
arrow
A prefix cache that only helps one process is useful, but requests move across replicas, HBM fills, sessions spill, and workers restart, so reuse that lives inside a single worker is not a fleet architecture. Part 2 works through what the serving stack needs once KV state leaves local GPU memory: a tier that is larger than HBM, fast enough that restore beats recompute, shared across workers, and reachable through the runtime's own KV transfer path, which is memory behavior at cluster scope rather than storage.
Agent Memory
AI/ML
AIStor
Performance
Prompt Caching Is an AI Margin Lever, Not a Model Trick | Part 1 of 3
arrow
Agentic AI applications resend the same project rules, tool schemas, and document context on every turn, so an expensive GPU fleet spends much of its time rebuilding a prefix it has already processed. Part 1 of three reframes prompt caching as an operating-margin lever rather than a model feature, maps prompt caching, prefix caching, KV cache, and KV cache offload to the business questions each one answers, and argues that reusable context needs a memory path rather than ordinary enterprise storage.
Agent Memory
AI/ML
AIStor
Performance
We deleted the agent mid-sentence. The work continued.
arrow
Worker A gathers evidence, publishes an accepted handoff, starts another edit, and is deleted mid-sentence with the unfinished tail left visible. Worker B starts in a fresh runtime with no session state, verifies the last accepted boundary in the same authorized AIStor Memory Workspace, discards the unchecked tail, and continues the work rather than restarting it.
Agent Memory
AI/ML
AIStor
Your inbox agent has no business remembering your workouts
arrow
Personal AI agents become useful as they learn you, but that familiarity should not require one agent accumulating your entire life. AIStor Memory gives each agent a bounded relationship with its own learned history, active workspace, and credential scope, all under your control.
AI/ML
Agent Memory
AIStor
Introducing AIStor Memory: Long-Term Memory For AI Agents
arrow
Every agent begins with the experience your organization has already earned.
AI/ML
AIStor
Integrations & Partners
Agent Memory
MINIO AI'STOR NVIDIA logos over colorful smoke and GPU chip on dark background.
Enterprise AI Infrastructure Made Easy with AIStor and NVIDIA GPUs
arrow
AIStor integrates NVIDIA GPU Operator to automate GPU setup & management for AI workloads in Kubernetes
AI/ML
AIStor
Minio AI Stor promptObject logo with colorful paint brush strokes on white background.
The most powerful S3 API ever? Introducing the Prompt API.
arrow
PromptObject API lets you "talk" to unstructured objects using natural language—AI-powered data interactions
AIStor
Logos of MinIO, Dremio (with narwhal icon), and Iceberg on a blue-green gradient background.
Query Iceberg Tables on MinIO with Dremio
arrow
Tutorial: Set up Dremio on Kubernetes to query Apache Iceberg tables stored on MinIO object storage
Integrations & Partners
Data Lakes & Analytics
Glowing orange and yellow sphere above a black square on a blue network grid background with text.
Putting a Filesystem on Top of an Object Store is a Bad Idea. Here is why.
arrow
Filesystem on object store is bad idea—POSIX translation kills performance, security & data integrity vs native S3 API
Architecture & Design Patterns
Storage & Infrastructure
Performance
Apache Spark logo above MinIO and Kafka logos on a dark red and purple gradient background.
End to End Spark Structured Streaming for Kafka Topics
arrow
Create Kafka events & consume to MinIO end-to-end with Spark Streaming—3hrs reduced to <10mins
Apache Ecosystem
Kubernetes & Containers
Cloud Infrastructure
White Apache Spark, MinIO, and Kafka logos on a purple and orange gradient background.
Spark Structured Streaming With Kafka and MinIO
arrow
Spark Structured Streaming tutorial processing Kafka events into MinIO with checkpointing
Apache Ecosystem
MinIO logo above Kafka and Kubernetes logos on a dark background with gear outlines.
How to Set up Kafka and Stream Data to MinIO in Kubernetes
arrow
Set up Kafka on Kubernetes & stream data to MinIO using Kafka Connect for real-time data lakes
Apache Ecosystem
Operations
Kubernetes & Containers
Logos of Minio, Dremio with a dolphin, and Kubernetes with a ship wheel on dark gradient background.
Dremio and MinIO on Kubernetes for Fast Scalable Analytics
arrow
Analytics on Kubernetes tutorial with Dremio SQL engine and MinIO lakehouse storage
Data Lakes & Analytics
Kubernetes & Containers
Integrations & Partners
Apache Ecosystem
MINIO and ICEBERG logos with geometric iceberg icon on a blue digital network background.
Manage Iceberg Tables with Spark
arrow
Tutorial for managing Apache Iceberg tables using Spark with MinIO object storage backend
Apache Ecosystem
Integrations & Partners
Architecture & Design Patterns
Storage & Infrastructure
Logos of MinIO, Apache Spark, and Kubernetes on a dark blue textured background with a plus sign.
Spark, MinIO and Kubernetes
arrow
Spark with MinIO on Kubernetes—deploy distributed Spark analytics on cloud-native object storage for lakehouse workloads
Apache Ecosystem
Kubernetes & Containers