#

prompt-compression

Here are 21 public repositories matching this topic...

aeromomo / claw-compactor

🦞 LLM Token Compression & Reduction Tool — Cut AI agent token costs by up to 97%. 6-layer deterministic context compression for AI agent workspaces. No LLM required. Prompt compression, context window optimization & cost reduction for any LLM pipeline.

Updated Mar 10, 2026
Python

atjsh / llmlingua-2-js

JavaScript/TypeScript implementation of LLMLingua-2 (Experimental)

nodejs javascript typescript web tensorflow transformers webgpu hf tensorflowjs prompt-engineering transformer-js prompt-compression llmlingua

Updated Sep 14, 2025
TypeScript

centminmod / or-cli

Python command-line tool for interacting with AI models through the OpenRouter API/Cloudflare AI Gateway, or local self-hosted Ollama. Optionally support Microsoft LLMLingua prompt token compression

openai linkup opik rag openai-api txtai llms llm-inference openrouter ollama cloudflare-ai ollama-api prompt-compression structured-outputs openai-api-client openrouter-api cloudflare-ai-gateway ai-rag llmlingua

Updated Dec 28, 2025

NodeNestor / claude-rolling-context

Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

claude ai-agent anthropic context-window context-management prompt-compression context-compression llm-context ai-coding claude-code claude-code-plugin claude-code-extension rolling-context

Updated Mar 10, 2026
Python

napmany / cutia

CUTIA: compress prompts while preserving quality

dspy prompt-engineering prompt-compression

Updated Feb 2, 2026
Python

kaistAI / GenPI

This repository is the official implementation of Generative Context Distillation.

agent distillation prompt-injection prompt-compression prompt-internalization context-distillation

Updated May 10, 2025
Python

therohanparmar / t3-toon

TOON for TYPO3 — a compact, human-readable, and token-efficient data format for AI prompts & LLM contexts. Perfect for ChatGPT, Gemini, Claude, Mistral, and OpenAI integrations (JSON ⇄ TOON).

Updated Mar 2, 2026
PHP

Kelpejol / prompt-compression-gateway

API gateway for LLM prompt compression with policy enforcement built on LLMLingua. Demonstrates cost control, prompt safety, and LLM execution boundaries.

python api-gateway fastapi llm prompt-compression

Updated Dec 26, 2025
Python

contextcrunch-ai / contextcrunch-python

Compress LLM Prompts and save 80%+ on GPT-4 in Python

python api llm prompt-compression

Updated Jan 17, 2024
Python

chirindaopensource / compact_prompt_unified_pipeline_prompt_data_compression_LLM_workflows

End-to-End Python implementation of CompactPrompt (Choi et al., 2025): a unified pipeline for LLM prompt and data compression. Features modular compression pipeline with dependency-driven phrase pruning, reversible n-gram encoding, K-means quantization, and embedding-based exemplar selection. Achieves 2-4x token reduction while preserving accuracy.

Updated Nov 30, 2025
Jupyter Notebook

ksm26 / Prompt-Compression-and-Query-Optimization

Enhance the performance and cost-efficiency of large-scale Retrieval Augmented Generation (RAG) applications. Learn to integrate vector search with traditional database operations and apply techniques like prefiltering, postfiltering, projection, and prompt compression.

Updated Jul 23, 2024
Jupyter Notebook

Starscream-11813 / Frugal-ICL

This repository contains the code and data of the paper titled "FrugalPrompt: Reducing Contextual Overhead in Large Language Models via Token Attribution."

prompt-compression frugal-ai token-attribution globenc decompx frugal-prompt

Updated Mar 6, 2026
Jupyter Notebook

sidedwards / tinyprompt

A fast, Unix-style CLI tool for semantic prompt compression. Cuts LLM prompt tokens by 10-20x with >90% fidelity, saving costs and latency.

cli text-processing compresssion llm llmops prompt-compression

Updated Sep 19, 2025
Python

ottobot2025 / SPEC-compression

Prompt compaction and shorthand codec for LLM workflows

text-compression llm prompt-compression

Updated Mar 9, 2026
Python

desagencydes-rgb / CATALYST

CATALYST - Lightning-fast optimization plugin for Claude Code + Ollama. Achieves 3-4x speedup through intelligent prompt compression, smart caching, and task-aware planning. Zero dependencies, MIT licensed, production-ready.

plugin caching performance optimization speed developer-tools local-models ollama prompt-compression claude-code

Updated Feb 28, 2026
JavaScript

SreeyaSrikanth / RL-Prompt-Compression

RL-Prompt-Compression employs graph-enhanced reinforcement learning with a Phi-3 compressor trained via GRPO using a TinyLlama evaluator and a MiniLM cross-encoder feedback model, to optimize prompt compression and improve model efficiency.

reinforcement-learning prompt-compression

Updated Nov 11, 2025
Jupyter Notebook

npow / kompact

LLM context compression proxy — 40-70% token savings, zero code changes

python proxy openai tfidf ai-agents claude fastapi gpt4 llm cost-reduction tiktoken anthropic context-window llm-optimization prompt-compression context-compression token-optimization

Updated Mar 9, 2026
Python

sriinnu / clipforge-PAKT

PAKT: Lossless prompt compression for LLMs. 30-50% fewer tokens on JSON/YAML/CSV/Markdown. Perfect round-trip fidelity. TypeScript library + CLI + Chrome extension + Tauri desktop app.

Updated Mar 8, 2026
TypeScript

npow / context-bench

Benchmark any system that transforms LLM context: compressors, RAG rerankers, memory managers, and more.

python nlp benchmark openai evaluation-framework rag openai-api ai-tools llm context-window llm-evaluation context-management prompt-compression token-optimization llm-benchmark

Updated Mar 9, 2026
Python

ofershap / prompt-compression

ai-agent context-window prompt-compression cursor-plugin claude-code token-optimization agents-md

Updated Feb 21, 2026
JavaScript

Improve this page

Add a description, image, and links to the prompt-compression topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the prompt-compression topic, visit your repo's landing page and select "manage topics."