GPU backends converge on sparse attention as vLLM wires DeepSeek V4.1 indexer
- OpenVINO: optimize stateful decode and GPU MoE inference ggml-org/llama.cpp
- OpenVINO: optimize stateful decode and GPU MoE inference (#28638) ggml-org/llama.cpp
$ git log --author="wine99" --all --oneline
Every morning, RepoJournal's newsroom reads what shipped across the open-source projects developers depend on — and writes the wire. Your commits kept making the news. This is your clippings file: the stories where your work was the story.
3
stories filed
2
daily wires
$ grep -rn "author: wine99" wires/ | sort -r
These clippings exist because your work kept showing up in other projects' wires. RepoJournal can write the whole journal: your commits, PRs and releases, turned into a daily devlog you never have to write.
Free. We read your public contributions only - private repos are never summarised.