Skip to content

Add deepseek-v4-flash-0731-in-c inference framework - #195

Open
shyringo wants to merge 1 commit into
xlite-dev:mainfrom
shyringo:add-deepseek-v4-flash-0731-in-c
Open

Add deepseek-v4-flash-0731-in-c inference framework#195
shyringo wants to merge 1 commit into
xlite-dev:mainfrom
shyringo:add-deepseek-v4-flash-0731-in-c

Conversation

@shyringo

Copy link
Copy Markdown

Adds deepseek-v4-flash-0731-in-c to LLM Train/Inference Framework/Design.

It is an Apache-2.0 C99 + OpenMP runtime that executes the native DeepSeek-V4-Flash-0731 Safetensors checkpoint on CPU-only machines without CUDA, PyTorch, or weight conversion. The runtime streams cold MoE experts from storage and documents its architecture, reproducible CPU measurements, correctness oracles, CI, and implementation provenance.

The change follows the table's existing code-project format and adds one README row with a documentation link, code link, and GitHub star badge.

Disclosure: I maintain the project being added.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant