AI SRE tools for RCA, Incident Response, Cost-Saving, Infra management, DevOps and more
-
Updated
Sep 1, 2026 - JavaScript
AI SRE tools for RCA, Incident Response, Cost-Saving, Infra management, DevOps and more
ultra-lightweight, mathematically robust prompt compression middleware
Laravel AI Guard 🛡️💰🤖 - Control and optimize AI costs in Laravel AI SDK applications 🚀 Track OpenAI & LLM token usage 📊, estimate AI costs before execution
Local, cache-aware LLM usage and cost telemetry for OpenClaw.
Production operations framework for AI-powered SaaS. The architectural patterns, failure modes, and operational playbooks that determine whether your AI systems scale profitably or fail expensively.
Token-efficient web research for AI agents; tinyfish search + Groq summarisation, 99% fewer tokens than raw HTML
An intelligent, low-latency local LLM router that reduces AI costs by 30-70%. Uses a self-hosted classifier to automatically route prompts to the most cost-effective model without external API overhead.
OpenAI-compatible LLM gateway that reduces API costs using Redis exact cache and Qdrant semantic cache.
AI Image Generation Cost Analysis
Rust CLI that reduces Claude Code token usage by 60-90%. Transparent proxy for git, find, grep, and dev commands — filters noise before it hits the context window.
Open-source, self-hostable AI cost & workflow observability. Find the prompt, customer, model, and workflow path behind every LLM cost spike — without a proxy.
System-level lint for multi-agent harnesses. Catches the 21 structural traps single-file linters miss — including the LLM-when-you-should-use-code patterns that burn tokens.
Local hooks that catch vague AI-agent prompts before they burn tokens.
Cut Claude Code spend without sacrificing quality — and prove it. Haiku/Sonnet/Opus router with real $-saved numbers, not vibes.
AI-powered AWS cost analysis and optimization agent using natural language — built with Amazon Bedrock AgentCore and Strands Agents SDK
OpenAI-compatible proxy enforcing SKILL.state (arXiv:2608.26263) — bounded O(1) prompts, O(T) total tokens for long-horizon LLM agents. Drop-in for Venice, OpenAI, Anthropic, OpenRouter.
The proxy that keeps your AI prompt cache warm — and your bill low. Anthropic · OpenAI · Gemini · Grok, one base-URL swap.
AI Image Generation Cost Calculator 2026 💰 - Compare & Save
Free read-only AI/LLM API cost review for GitHub: find token, retry, cache, and model-call signals without uploading source.
# AWS Bedrock Claude REST API with Terraform This project provides a complete Terraform setup to expose Claude AI models through AWS Bedrock via a REST API. All usage is billed directly through AWS, eliminating the need for separate Anthropic API credits.
To associate your repository with the ai-cost-optimization topic, visit your repo's landing page and select "manage topics."