跳到主要内容 / Skip to main content
GitHub819github.com

lyogavin/airllm

by lyogavin首次上榜 2026-07-19

摘要

AirLLM 能在单块 4GB 显存 GPU 上运行 70B 参数的大模型推理,不需要多卡或高显存设备。它显著降低了大模型本地部署的硬件门槛,适合资源有限的开发者尝试超大参数模型。

相关推荐

TencentCloud/TencentDB-Agent-Memory

TencentDB Agent Memory is a team-level memory hub for AI Agents — turning conversations, docs, and code into four reusable memory assets (Chat Memory, Skill, LLM-Wiki, Code-Graph) that are governed, shared, and equipped across agents and frameworks.

agentaiassetchatconversation

diegosouzapw/OmniRoute

Never stop coding. Free MIT AI gateway: one endpoint, 290+ providers (90+ free), 500+ models — Kimi, Claude, GPT, OpenAI, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 500+ contributors

aiautoawarecavemanclaude

koala73/worldmonitor

Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastructure tracking in a unified situational awareness interface

aiawarenessdashboardgithubinfrastructure