AETHER · Blog

More efficient AI systemsTechnical notes on.

Research and application papers on context architecture, token cost, RAG quality and LLM operations in production environment.

Latest Posts

Research and application notes.

Contents on measurement methods, technical decisions and production experiences.

Context Engineering01
CONTEXT / COST

Why doesn't cost remain linear as the LLM context grows?

We examine how token cost accumulates when long conversations, system instructions, and repetitive product context are carried over again with each request.

12 August 2026 · 7 min read
Read the article
RAG Systems02
RETRIEVAL / SIGNAL

Is more documentation always better in RAG pipelines?

Quality and cost implications of selecting the context associated with the query rather than sending entire chunks returned from the retrieval layer to the model.

August 5, 2026 · 9 min read
Read the article