Model capability radar plugin for the DeepSeek Harness Web GUI
-
Updated
Sep 14, 2026 - TypeScript
Model capability radar plugin for the DeepSeek Harness Web GUI
Public report and aggregate evidence for LayerRail's July 2026 frontier-model deployment study
This script cleans passenger data, creates family features, tests eight models including XGBoost and CatBoost, and combines their votes to predict survival.
Native macOS menu-bar watcher for AI Stupid Level: top 20, value rankings, GPT/Claude clusters, C/R/T signals, and a notch island.
Tariff comparison tool for Mistral Large 4 users — track usage limits across plans.
End-to-end Deep Learning Vision Benchmark evaluating 11 architectures across 5 paradigms. Features Soft-Voting Ensemble (98.92% Top-5), ONNX INT8 Quantization (3.01x speedup), Grad-CAM XAI, FastAPI microservice, and live Streamlit Arena.
Healthcare demand forecasting and staffing decision platform (PoC): a 13-model benchmark (baselines, SARIMAX, Prophet, global gradient boosting, Nixtla Stats/ML/Neural, Chronos), conformal prediction intervals, rolling-origin backtesting, a costed staffing layer, batch pipeline, FastAPI serving, monitoring and responsible-ML docs.
Ask one question to multiple LLMs simultaneously — auto-discovers configured models, parallel calls, side-by-side comparison
To associate your repository with the model-benchmark topic, visit your repo's landing page and select "manage topics."