Skip to content
 
 

Latest commit

 

History

477 Commits

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

Daily arXiv AI Enhanced · Self-hosted

Track arXiv daily. Summarize, rank and review papers with your own Claude Code / Codex CLI — no third-party API key required.

简体中文 · English · Full Guide


Paper list

What this is

Upstream dw-dengwei/daily-arXiv-ai-enhanced runs on GitHub Actions with a third-party LLM API key. This fork rewires it to local cron + local CLI: it drives the Claude Code or Codex CLI you have already logged into, using your own subscription quota. No API key to obtain, and every paper and review stays on your machine.

Upstream This fork
Runtime GitHub Actions local cron
Model DeepSeek / OpenAI API key your logged-in Claude Code or Codex
Cost pay per token your existing subscription, no API key
Frontend multi-page single-page app, no reload between views
Ordering by category ranked 0–10 against your research focus
Extras — one-click review, cross-device bookmarks, trend charts, dark mode

Features

Relevance ranking, not a category dump

Write your research directions in ai/research_focus.txt — it is the sole basis for scoring. Every paper gets a 0–10 score and a full five-section summary from the fast model; those above the threshold and in the day's top-K get rewritten by the strongest model. Edit the file and the next run picks it up — no code changes.

One-click review

One-click review

Two stages: a fast model picks the most likely venue from a fixed whitelist (NDSS / USENIX Security / ICSE / NeurIPS …), then Quick / Normal / Deep selects the model that writes the review as a reviewer for that venue — summary, strengths, weaknesses, detailed comments, questions to the authors, rating, recommendation and confidence.

The tiers genuinely differ. On the same paper the fast model returned "Minor revision, 8/10" while the mid tier returned "Major revision, 4/10" and questioned a specific latency claim against the typical latency of comparable policy engines.

Trends

Trends

Daily trends for primary categories, sub-directions and keywords. A crosshair reads out every series at once, the legend toggles series on and off, and a table view is one click away. Sub-direction labels are bootstrapped: the model reuses existing labels and only creates a new one when nothing fits.

Bookmarks × review verdicts

Bookmarks

Bookmarks and reviews live in one sqlite file, so the bookmark list shows the recommendation, rating and venue inline instead of making you open each paper. There is a "reviewed only" filter.

Interface language

Language picker on first run

On first run the app asks which language you want, and remembers it. Change it any time under Settings → Language. The choice drives three things, which take effect differently:

What When it changes
Interface text immediately
New reviews the next review you run is written in that language
Paper summaries fixed when the daily pipeline generated them

Summaries are the one that cannot change retroactively — they were written by the pipeline at crawl time. If the language you picked has no summaries for the range you are looking at, the app says so and tells you which LANGUAGE to set in .env.local. Existing reviews are kept in the language they were written in rather than being re-translated.

Dark mode · works on phones

Light Mobile

Black/white/green palette, following the system or toggled manually. The theme is applied before first paint, so dark-mode users never see a white flash.

Quick start

# 1. Install and log into one CLI
npm i -g @anthropic-ai/claude-code && claude     # https://claude.com/claude-code
# or
npm i -g @openai/codex && codex login            # https://developers.openai.com/codex/cli

# 2. Install
git clone https://github.com/springkill/daily-arXiv-ai-enhanced.git
cd daily-arXiv-ai-enhanced
python3 -m venv .venv && source .venv/bin/activate && pip install -e .

# 3. Configure
cp .env.local.example .env.local && $EDITOR .env.local
cp ai/research_focus.example.txt ai/research_focus.txt && $EDITOR ai/research_focus.txt
python3 ai/llm.py "reply with exactly: OK"       # smoke-test the CLI

# 4. Run
./run-local.sh
python3 -m http.server 8000                      # http://localhost:8000

Bookmarks and reviews need the backend (optional — everything else works without it):

ln -sf "$PWD/deploy/api/arxiv-api.service" ~/.config/systemd/user/
systemctl --user daemon-reload && systemctl --user enable --now arxiv-api

For production deployment see SELF-HOSTED.md; for how it works and every config knob see docs/GUIDE.md.

Where your data lives

data/ (daily jsonl), var/store.sqlite3 (bookmarks + reviews), .env.local, ai/research_focus.txt and assets/trend-taxonomy.json are all gitignored — the repository ships only .example / .seed templates and contains nobody's personal data.

Caution

If your jurisdiction has censorship requirements for academic data, run this code with caution; any redistributed version must fulfil its content review obligations (including but not limited to the compliance of the original papers and of AI output), otherwise all legal consequences are borne by the downstream party.


Contributors

Most of this project's functionality comes from upstream dw-dengwei/daily-arXiv-ai-enhanced. Thanks to the following contributors of the original project for contributing code, discovering bugs, and sharing useful ideas:

JianGuanTHU
JianGuanTHU

Chi-hong22
Chi-hong22

chaozg
chaozg

quantum-ctrl
quantum-ctrl

Zhao2z
Zhao2z

eclipse0922
eclipse0922

xuemian168
xuemian168

Lrrrr549
Lrrrr549

AinzRimuru
AinzRimuru

fengxueguiren
fengxueguiren

fengxueguiren
zerocpp

Acknowledgement

Sincere thanks to the following individuals and organizations for promoting and supporting the original project:

Github_Daily
Github_Daily

AIGCLINK
AIGCLINK

阮一峰的网络日志
阮一峰的网络日志
科技爱好者周刊
(第 353 期)

《HelloGitHub》第 111 期
《HelloGitHub》
月刊第 111 期

Upstream

This fork changes the runtime and the frontend; the core idea, the crawler and the data format all come from upstream.

License

Same license as upstream, see LICENSE.

About

Automatically crawl arXiv papers daily and summarize them using AI. Illustrating them using GitHub Pages.

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages