Track arXiv daily. Summarize, rank and review papers with your own Claude Code / Codex CLI — no third-party API key required.
简体中文 · English · Full Guide
Upstream dw-dengwei/daily-arXiv-ai-enhanced runs on GitHub Actions with a third-party LLM API key. This fork rewires it to local cron + local CLI: it drives the Claude Code or Codex CLI you have already logged into, using your own subscription quota. No API key to obtain, and every paper and review stays on your machine.
| Upstream | This fork | |
|---|---|---|
| Runtime | GitHub Actions | local cron |
| Model | DeepSeek / OpenAI API key | your logged-in Claude Code or Codex |
| Cost | pay per token | your existing subscription, no API key |
| Frontend | multi-page | single-page app, no reload between views |
| Ordering | by category | ranked 0–10 against your research focus |
| Extras | — | one-click review, cross-device bookmarks, trend charts, dark mode |
Write your research directions in ai/research_focus.txt — it is the sole basis
for scoring. Every paper gets a 0–10 score and a full five-section summary from the
fast model; those above the threshold and in the day's top-K get rewritten by the
strongest model. Edit the file and the next run picks it up — no code changes.
Two stages: a fast model picks the most likely venue from a fixed whitelist (NDSS / USENIX Security / ICSE / NeurIPS …), then Quick / Normal / Deep selects the model that writes the review as a reviewer for that venue — summary, strengths, weaknesses, detailed comments, questions to the authors, rating, recommendation and confidence.
The tiers genuinely differ. On the same paper the fast model returned "Minor revision, 8/10" while the mid tier returned "Major revision, 4/10" and questioned a specific latency claim against the typical latency of comparable policy engines.
Daily trends for primary categories, sub-directions and keywords. A crosshair reads out every series at once, the legend toggles series on and off, and a table view is one click away. Sub-direction labels are bootstrapped: the model reuses existing labels and only creates a new one when nothing fits.
Bookmarks and reviews live in one sqlite file, so the bookmark list shows the recommendation, rating and venue inline instead of making you open each paper. There is a "reviewed only" filter.
On first run the app asks which language you want, and remembers it. Change it any time under Settings → Language. The choice drives three things, which take effect differently:
| What | When it changes |
|---|---|
| Interface text | immediately |
| New reviews | the next review you run is written in that language |
| Paper summaries | fixed when the daily pipeline generated them |
Summaries are the one that cannot change retroactively — they were written by the
pipeline at crawl time. If the language you picked has no summaries for the range
you are looking at, the app says so and tells you which LANGUAGE to set in
.env.local. Existing reviews are kept in the language they were written in
rather than being re-translated.
Black/white/green palette, following the system or toggled manually. The theme is applied before first paint, so dark-mode users never see a white flash.
# 1. Install and log into one CLI
npm i -g @anthropic-ai/claude-code && claude # https://claude.com/claude-code
# or
npm i -g @openai/codex && codex login # https://developers.openai.com/codex/cli
# 2. Install
git clone https://github.com/springkill/daily-arXiv-ai-enhanced.git
cd daily-arXiv-ai-enhanced
python3 -m venv .venv && source .venv/bin/activate && pip install -e .
# 3. Configure
cp .env.local.example .env.local && $EDITOR .env.local
cp ai/research_focus.example.txt ai/research_focus.txt && $EDITOR ai/research_focus.txt
python3 ai/llm.py "reply with exactly: OK" # smoke-test the CLI
# 4. Run
./run-local.sh
python3 -m http.server 8000 # http://localhost:8000Bookmarks and reviews need the backend (optional — everything else works without it):
ln -sf "$PWD/deploy/api/arxiv-api.service" ~/.config/systemd/user/
systemctl --user daemon-reload && systemctl --user enable --now arxiv-apiFor production deployment see SELF-HOSTED.md; for how it works and every config knob see docs/GUIDE.md.
data/ (daily jsonl), var/store.sqlite3 (bookmarks + reviews), .env.local,
ai/research_focus.txt and assets/trend-taxonomy.json are all gitignored —
the repository ships only .example / .seed templates and contains nobody's
personal data.
Caution
If your jurisdiction has censorship requirements for academic data, run this code with caution; any redistributed version must fulfil its content review obligations (including but not limited to the compliance of the original papers and of AI output), otherwise all legal consequences are borne by the downstream party.
Most of this project's functionality comes from upstream dw-dengwei/daily-arXiv-ai-enhanced. Thanks to the following contributors of the original project for contributing code, discovering bugs, and sharing useful ideas:
|
JianGuanTHU |
Chi-hong22 |
chaozg |
quantum-ctrl |
Zhao2z |
eclipse0922 |
|
xuemian168 |
Lrrrr549 |
AinzRimuru |
fengxueguiren |
zerocpp |
Sincere thanks to the following individuals and organizations for promoting and supporting the original project:
![]() Github_Daily |
![]() AIGCLINK |
阮一峰的网络日志 科技爱好者周刊 (第 353 期) |
![]() 《HelloGitHub》 月刊第 111 期 |
- Original: dw-dengwei/daily-arXiv-ai-enhanced
- Roadmap: https://github.com/users/dw-dengwei/projects/3
This fork changes the runtime and the frontend; the core idea, the crawler and the data format all come from upstream.
Same license as upstream, see LICENSE.









