Bookmarks
Tag cloud
Picture wall
Daily
Search
RSS Feed
RSS Feed
Daily Feed
Weekly Feed
Monthly Feed
Filters
Links per page
20 links
50 links
100 links
Custom value
Filters
Untagged links
tags
search
1 result tagged
amd
✕
GitHub - peonist-ai/halogen-flash-server: The fastest way to run Qwen3.8-Flash-Next on Strix Halo (gfx1151) · GitHub
https://github.com/peonist-ai/halogen-flash-server
Wed Sep 16 08:32:44 2026
GitHub - rasbt/LLMs-from-scratch: Implement a ChatGPT-like LLM in PyTorch from scratch, step by step · GitHub
: Implement a ChatGPT-like LLM in PyTorch from scratch, step by step - rasbt/LLMs-from-scratch
GitHub - Mega4alik/ollm
: Contribute to Mega4alik/ollm development by creating an account on GitHub.
av/awesome-llm-services: A list of self-hostable LLM services
:
GitHub - prayangshuuu/hummingbird: hummingbird is a lightweight, zero dependency runtime for massive open source Mixture of Experts (MoE) language models. It unifies SSD, RAM, and VRAM into a single intelligent memory hierarchy, enabling inference of models like GPT-OSS 120B, GLM, DeepSeek, Qwen, and more on consumer hardware · GitHub
:
GitHub - nexu-io/open-design: 🎨 Local-first, open-source alternative to Anthropic's Claude Design. ⚡ 19 Skills · ✨ 71 brand-grade Design Systems 🖼 Generate web · desktop · mobile prototypes · slides · images · videos · HyperFrames 📦 Sandboxed preview · HTML/PDF/PPTX/MP4 export 🤖 Runs on Claude Code / Codex / Cursor / Gemini / OpenCode / Qwen / Copilot / Hermes / Kimi CLI. · GitHub
: 🎨 Local-first, open-source alternative to Anthropic's Claude Design. ⚡ 19 Skills · ✨ 71 brand-grade Design Systems 🖼 Generate web · desktop · mobi...
The fastest way to run Qwen3.8-Flash-Next on Strix Halo (gfx1151) - peonist-ai/halogen-flash-server
2879 links, including 104 private