Code scanner to check for issues in prompts and LLM calls
-
Updated
Apr 6, 2025 - Python
Code scanner to check for issues in prompts and LLM calls
Turbocharged TensorFlow fork with experimental TurboQuant extension. High-performance weight-only, block-wise codebook quantization for Keras layers.
Building an AI team to play Codenames using top Large Language Models (LLMs), evaluating performance, and pitting them against each other. Explore their strategy and capabilities in this interactive competition!
La Perf is a framework for AI performance benchmarking — covering LLMs, VLMs, embeddings, with power-metrics collection.
Arbitrary Numbers
KAI Data Center Builder
Boost FPS 2026: Free AI Optimizer Tool ⚡ - One-Click PC Boost
Powerful AI efficiency tool that reduces token usage by up to 75% for cloud code and LLM applications. Ideal for developers looking to maximize performance while minimizing costs in 2026.
AI Performance Engineering Cheatsheet: From Cloud to Edge.
Test AI provider latency (TTFB, TTFT, TPS) in your CI/CD pipeline. Benchmark OpenAI, Anthropic, Google, and more.
Chrome extension that removes old ChatGPT messages from the DOM to keep long conversations fast and responsive.
Correctness-first microbenchmarks for LLM attention and sampling kernels.
A streamlined and easy-to-use AI performance evaluation / summary template with modern UI in HTML, including correct percentage chart and comparison with other models, precision, recall, F1-score, and confusion matrix. Enables you to create the result chart within 3 minutes.
Speedtest for AI. Test latency to every major AI provider from your terminal.
Add a description, image, and links to the ai-performance topic page so that developers can more easily learn about it.
To associate your repository with the ai-performance topic, visit your repo's landing page and select "manage topics."