芦苇
发帖子
探索
今天发现最新帖子添加订阅
热门板块
🤖AI💡科技💻开发🧭产品🛠️工具
订阅
关注
下载芦苇 App ↗
我我的

NVIDIA dropped an NVIDIA-hosted CUDA MCP for AI-assisted CUDA operations, such as searching official, up-to-date documentation, writing optimized GPU code, and analyzing performance data

localllama (Reddit),r/LocalLLaMA,约 72.8 万成员的社区,专注于在本地硬件上运行大语言模型,涵盖模型选择、GPU 配置、量化技术、Ollama 与 llama.cpp 等工具,以及隐私优先的 AI 工作流。关注
添加评论
点赞收藏
点踩分享查看原文
评论
?
参与讨论

登录芦苇

登录后关注作者、收藏内容和参与讨论。

关于作者
localllama (Reddit)r/LocalLLaMA,约 72.8 万成员的社区,专注于在本地硬件上运行大语言模型,涵盖模型选择、GPU 配置、量化技术、Ollama 与 llama.cpp 等工具,以及隐私优先的 AI 工作流。
相关文章
Fine-tuning Cactus Needle 2 can match DeepSeek v4 on the specific task查看相关内容
I pushed Qwen3.8-27B to 381 tps for a single request on a RTX 3090查看相关内容
If you are wondering why Ornith 1.5 35B A3B with MTP is so slow, this is why查看相关内容