Spaces:
Build error
Build error
Commit History
Add OSS evaluation suite, universal JSON crush, latency benchmarks ad6f44c
Merge pull request #33 from angpt/angpt/any-llm-integration 713e4d6
Update CHANGELOG.md adf7061
Bump version to 0.3.7 f5eea12
Fix CI: guard starlette imports, asyncio.run(), deprecate datetime.utcnow() 5f8e891
Fix: guard starlette imports in test_compress_api.py for CI without proxy deps 2a0de14
README: replace plain text architecture with Mermaid diagrams e5bafa9
README: add back proof points — needle-in-haystack demo, benchmarks, architecture detail ff29454
README: proxy as the hero quickstart, compress() for Python, integration table for existing proxies 0fb4bed
Rewrite README + add Integration Guide 81cb88d
Add one-function compress() API, ASGI middleware, LiteLLM callback 6505c43
Add Compression Hooks — extension points for SaaS and advanced customization 6aca3c8
Update anyllm.py 85a32a8
Bump version to 0.3.6 4a263c2
Add Query Echo: re-inject user question after compressed tool outputs 751fb9e
Add compression summaries, multi-provider headers, Dockerfile fix 4d7e460
Fixing the DockerFile to have hnswlib working 3a888ea
Add pluggable adapter hooks for CCR, Storage, and TOIN backends 107515b
Add tests for streaming resilience and concurrent session safety 83c709c
Fix streaming crashes under concurrent sessions 3b0b075
Update anyllm.py 141dfaf
Bump version to 0.3.2 419f040
Add adaptive compression sizing and fix cost tracking bugs 9efd75a
Update anyllm.py 7725ae9
Adding any-llm as a backed provider c106272
Fix security vulnerabilities in memory and CCR systems 3b38650
Add LiteLLM backend routing for OpenAI endpoint and Magika content detection 24e92f8
Fix token savings metrics mismatch between pipeline and server 0c64bc4
Fix inflated token savings stats from cache hits 66cdf98
Use AWS API to dynamically fetch Bedrock inference profiles 1e88e9a
Fix MCP tests to work when MCP SDK is not installed d45db57
Add MCP CLI for Claude Code subscription users bc5c41c
Merge pull request #26 from prakersh/dev/prakersh/fix 99b7457
fix: Handle BaseModelOutputWithPooling from SigLIP get_*_features() 8f2a8a6
Prakersh Maheshwari Claude Opus 4.5 commited on
Merge pull request #25 from prakersh/fix/increase-request-body-size-limit 25f316a
test: Update LLMLingua default model test to match new default ba757e5
Prakersh Maheshwari Claude Sonnet 4.5 commited on
fix: Increase MAX_REQUEST_BODY_SIZE from 10MB to 100MB to support image-heavy requests f56c6c9
Prakersh Maheshwari Claude Sonnet 4.5 commited on