macOS menubar app for fast local DeepSeek V4 Pro/Flash, with 1M context.
-
Updated
Aug 16, 2026 - Swift
macOS menubar app for fast local DeepSeek V4 Pro/Flash, with 1M context.
⚡️ A community driven PHP client for DeepSeek AI, designed to bring clean API access, fluent developer experience, and framework-friendly integration to PHP applications.
Fixes missing reasoning_content for DeepSeek V4
基于 FastAPI 的 DeepSeek Chat 反向代理,将 DeepSeek 网页版的 API 转换为 OpenAI 兼容格式。 支持流式/非流式对话、专家模式、深度思考(reasoning_content)、工具调用(DSML prompt injection)。 自动处理 PoW 鉴权挑战,无需官方 API Key。
DeepSeek V4 Flash CPU/NVMe research fork: 78.62 GiB GGUF validated on 7.7 GiB RAM, CPU-only, using demand paging.
DeepSeek-V4-Flash-0731 284B inference in ~30 GB of RAM on any M-series MacBook
A tool to have multiple claude-code instance with deepseek, minimax, and z.ai glm models
Kimi K3, deepseek v4 flash, glm 5.2, qwen 3.8 27b, Muse Glimmer, Nemotron, glm 5.3
DeepSeek 逆向 API 支持 Deepseek V4
Guide for development with DeepSeek Harness. Building plugin for DeepSeek Harness Project.
An easy-to-use, reliable DeepSeek Harness desktop app with robust plugin management, process orchestration, dynamic tool discovery, and a selection of useful pre-installed plugins that can be removed at any time.
Reproducible kit to deploy DeepSeek-V4-Flash-DSpark on a 2× NVIDIA DGX Spark (GB10) cluster: vLLM TP=2 over QSFP 200GbE, NVFP4 KV, DSpark speculative decoding, 1M context, systemd self-heal. Apache-2.0.
A collection of recipes/notebooks showcasing use-cases of open-source models with Qubrid AI.
Codex vision bridge for DeepSeek V4 Flash: give text-only DeepSeek image capability in Codex. Local proxy turns pasted images and view_image into text via free GLM-4V-Flash or any OpenAI-compatible vision API. No GPU, no Ollama.
DeepSeek-V4-Flash-0731 on 8x RTX 3090 (SM86) with vLLM — verified FP8 serving, benchmarks, build guide, and reproducible release.
Supercharge your daily workflows! Autonomously browse pages, scrape content, switch tabs, fill out inputs, and stream Chain-of-Thought (CoT) reasoning under granular human-in-the-loop safety switches and glassmorphic UI controls.
A Rust proxy that exposes an **OpenAI Responses API** interface, translating to various upstream backends (DeepSeek, OpenAI, Anthropic) with full streaming SSE support. Built for Codex CLI and similar tools hardcoded to OpenAI.
Unlock Claude Code with DeepSeek V4. Get Anthropic's agent tools with 95% lower costs and local vision.
Add a description, image, and links to the deepseek-v4-flash topic page so that developers can more easily learn about it.
To associate your repository with the deepseek-v4-flash topic, visit your repo's landing page and select "manage topics."