#
webbench
Here are 15 public repositories matching this topic...
Linux下高性能服务器,详细梳理后台各个知识点
-
Updated
Jun 7, 2020 - C++
Build an LLM inference stack from scratch in Rust with this free textbook covering kernels, KV caches, schedulers, and multi-GPU systems.
rust ai computer-vision book gpu webserver inference pytorch tokens language-model systems-programming swarm-intelligence webbench finetuning generative-ai instruction-tuning
-
Updated
Sep 12, 2026
Deploy a quantized Qwen3.8 27B coding agent locally on a 16GB GPU with llama.cpp, supporting OpenAI-compatible clients and optional Anthropic adapters for real-world repository tasks.
windows macos docker automation downloader csv gpu webserver discord-bot chinese all-in-one webbench ai-agent large-language-models flash-attention qwen vibe-coding
-
Updated
Sep 13, 2026 - Python
Optimize local LLM inference and benchmarking recipes for RTX 5060 Ti hardware configurations.
mysql swift database mongodb scale cpp webserver vue-router voice-chat distributed-transactions cloud-native agora model-deployment webbench model-monitoring clubhouse-lib tinywebserver mysql-compatibility clubhouse-client agent-context
-
Updated
Sep 12, 2026 - HTML
Add this topic to your repo
To associate your repository with the webbench topic, visit your repo's landing page and select "manage topics."