prompt-compression

Here are 9 public repositories matching this topic...

atjsh / llmlingua-2-js

JavaScript/TypeScript implementation of LLMLingua-2 (Experimental)

nodejs javascript typescript web tensorflow transformers webgpu hf tensorflowjs prompt-engineering transformer-js prompt-compression llmlingua

Updated Sep 14, 2025
TypeScript

centminmod / or-cli

Sponsor

Star

Python command-line tool for interacting with AI models through the OpenRouter API/Cloudflare AI Gateway, or local self-hosted Ollama. Optionally support Microsoft LLMLingua prompt token compression

openai linkup opik rag openai-api txtai llms llm-inference openrouter ollama cloudflare-ai ollama-api prompt-compression structured-outputs openai-api-client openrouter-api cloudflare-ai-gateway ai-rag llmlingua

Updated Aug 7, 2025

kaistAI / GenPI

Star

This repository is the official implementation of Generative Context Distillation.

agent distillation prompt-injection prompt-compression prompt-internalization context-distillation

Updated May 10, 2025
Python

contextcrunch-ai / contextcrunch-python

Star

Compress LLM Prompts and save 80%+ on GPT-4 in Python

python api llm prompt-compression

Updated Jan 17, 2024
Python

sidedwards / tinyprompt

Star

A fast, Unix-style CLI tool for semantic prompt compression. Cuts LLM prompt tokens by 10-20x with >90% fidelity, saving costs and latency.

cli text-processing compresssion llm llmops prompt-compression

Updated Sep 19, 2025
Python

ksm26 / Prompt-Compression-and-Query-Optimization

Star

Enhance the performance and cost-efficiency of large-scale Retrieval Augmented Generation (RAG) applications. Learn to integrate vector search with traditional database operations and apply techniques like prefiltering, postfiltering, projection, and prompt compression.

Updated Jul 23, 2024
Jupyter Notebook

Starscream-11813 / Frugal-ICL

Star

This repository contains the code and data of the paper titled "FrugalPrompt: Reducing Contextual Overhead in Large Language Models via Token Attribution."

prompt-compression frugal-ai token-attribution globenc decompx frugal-prompt

Updated Oct 30, 2025
Jupyter Notebook

chirindaopensource / compact_prompt_unified_pipeline_prompt_data_compression_LLM_workflows

Star

End-to-End Python implementation of CompactPrompt (Choi et al., 2025): a unified pipeline for LLM prompt and data compression. Features modular compression pipeline with dependency-driven phrase pruning, reversible n-gram encoding, K-means quantization, and embedding-based exemplar selection. Achieves 2-4x token reduction while preserving accuracy.

Updated Nov 30, 2025
Jupyter Notebook

SreeyaSrikanth / RL-Prompt-Compression

Star

RL-Prompt-Compression employs graph-enhanced reinforcement learning with a Phi-3 compressor trained via GRPO using a TinyLlama evaluator and a MiniLM cross-encoder feedback model, to optimize prompt compression and improve model efficiency.

reinforcement-learning prompt-compression

Updated Nov 11, 2025
Jupyter Notebook

Improve this page

Add a description, image, and links to the prompt-compression topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the prompt-compression topic, visit your repo's landing page and select "manage topics."

Learn more

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

prompt-compression

Here are 9 public repositories matching this topic...

atjsh / llmlingua-2-js

centminmod / or-cli

kaistAI / GenPI

contextcrunch-ai / contextcrunch-python

sidedwards / tinyprompt

ksm26 / Prompt-Compression-and-Query-Optimization

Starscream-11813 / Frugal-ICL

chirindaopensource / compact_prompt_unified_pipeline_prompt_data_compression_LLM_workflows

SreeyaSrikanth / RL-Prompt-Compression

Improve this page

Add this topic to your repo