Research Radar is an automated sweep of new work in AI/LLM safety, alignment, and pragmatic mechanistic interpretability. Claude Code cloud routines search OpenReview, the ACL Anthology, TMLR, arXiv (cs.CL / cs.LG / cs.CR / cs.AI), and the alignment forums, then write up what they find: a daily top ten Tuesday through Sunday, and a wider aggregate on Monday mornings.
Entries are ranked by importance to that agenda rather than by recency, and peer-reviewed work outranks incremental preprints. Each one carries a technical summary of method and result rather than a reprinted abstract, plus the venue and its peer-review status. If a day turns up fewer than ten genuinely relevant papers, the report says so instead of padding.
Window: August 5–7, 2026 · Sources swept: OpenReview, ACL Anthology, TMLR, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Mechanistic Interpretability Workshop · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: August 5–6, 2026 · Sources swept: OpenReview, ACL Anthology, TMLR, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), LessWrong/Alignment Forum, lab blogs · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: August 2–4, 2026 · Sources swept: OpenReview, ACL Anthology, TMLR, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), LessWrong/Alignment Forum, lab blogs · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: July 31 – August 2, 2026 · Sources swept: OpenReview, ACL Anthology, TMLR, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), LessWrong/Alignment Forum, lab blogs · Counts: 0 peer-reviewed · 10 preprints · 0 forum/blog
Window: July 30–August 1, 2026 (latest preprints); peer-reviewed sweep back through May 2026 for work not previously covered · Sources swept: OpenReview, ACL Anthology, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, GitHub · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: July 29–31, 2026 (latest preprints); back-swept for peer-reviewed work not previously surfaced · Sources swept: OpenReview, ACL Anthology, ESORICS proceedings, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: July 27–29, 2026 (latest preprints); back-swept through April 2026 for peer-reviewed work not previously surfaced · Sources swept: OpenReview, ACL Anthology, NeurIPS/ICLR proceedings, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar · Counts: 2 peer-reviewed · 1 workshop paper · 7 preprints · 0 forum/blog
Window: July 26–28, 2026 (latest preprints); back-swept through April 2026 for field-relevant work not previously surfaced · Sources swept: OpenReview, ACL Anthology, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, GitHub · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: July 24–26, 2026 (latest preprints); back-swept through April 2026 for field-relevant work not previously surfaced · Sources swept: OpenReview, ACL Anthology, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, GitHub · Counts: 0 peer-reviewed · 10 preprints · 0 forum/blog
Window: July 23–25, 2026 (latest preprints); back-swept through April 2026 for field-relevant work not previously surfaced · Sources swept: OpenReview, ACL Anthology, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, GitHub · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: July 21–24, 2026 (latest preprints); back-swept through May 2026 for field-relevant work not previously surfaced · Sources swept: OpenReview, ACL Anthology, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, LessWrong/Alignment Forum · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: July 22–23, 2026 (latest preprints); back-swept through June 2026 for field-relevant work not previously surfaced · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, GitHub · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: July 20–22, 2026 (latest preprints); back-swept through May 2026 for field-relevant work not previously surfaced · Counts: 0 peer-reviewed · 10 preprints · 0 forum/blog
Window: July 19–21, 2026 (sweeping back ~2 weeks for field-relevant work) · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, HuggingFace Papers, GitHub · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: July 13–22, 2026 · Fresh sweep covers: ICLR 2026 (OpenReview), ICML 2026 (main + Mechanistic Interpretability Workshop), ACL 2026 (main + Findings), NeurIPS 2025, COLM 2025, NeurIPS 2026 Competition Track, arXiv cs.CL/cs.LG/cs.CR/cs.AI · Counts: 8 peer-reviewed (3 ICLR 2026 · 1 NeurIPS 2025 · 1 ICML 2026 Mech Interp Workshop · 3 ACL 2026 main) · 7 preprints · 0 forum/blog
Window: 2026-07-17 to 2026-07-19 (new arXiv submissions); also surfaces ACL 2026 main + Findings papers (Jul 2–7) not covered in prior reports · Sources swept: arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ACL 2026 main, ACL 2026 Findings · Counts: 8 peer-reviewed · 2 preprints · 0 forum/blog
Window: 2026-07-16 to 2026-07-18 (new arXiv submissions); also surfaces high-relevance items from Jul 1–15 not covered in prior reports · Sources swept: arXiv (cs.CL/cs.LG/cs.CR/cs.AI), COLM 2025, ICML 2026 Mech Interp Workshop backfill · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: 2026-07-15 to 2026-07-17 (new arXiv submissions); also surfaces high-relevance items from Jul 1–14 not covered in prior reports · Sources swept: arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Mechanistic Interpretability Workshop, ACL 2026 Findings, NeurIPS 2026 Competition Track · Counts: 3 peer-reviewed · 7 preprints · 0 forum/blog
Window: 2026-07-14 to 2026-07-16 · Sources swept: OpenReview, ACL Anthology, IJCAI-ECAI 2026, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Mech Interp Workshop · Counts: 1 peer-reviewed · 5 preprints · 1 workshop spotlight (ICML 2026 Mech Interp)
Window: 2026-07-12 to 2026-07-15 (new preprints); broader backfill sweep for high-relevance items not covered in Jul 1–11 reports · Sources swept: arXiv (cs.CL/cs.LG/cs.CR/cs.AI), OpenReview (ICLR 2026, ICML 2026 Workshop on Mech Interp), Hugging Face Papers, Semantic Scholar · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: 2026-07-13 to 2026-07-14 (preprints last 48h); peer-reviewed items from ICLR 2026, ICML 2026, and EACL 2026 not previously covered; broader sweep of high-relevance preprints from Aug 2025–Jun 2026 not covered in prior reports · Sources swept: ICLR 2026 (OpenReview), ICML 2026 (main + AIWILD Workshop), EACL 2026 (ACL Anthology), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar · Counts: 5 peer-reviewed · 5 preprints · 0 forum/blog
Window: July 6–12, 2026 · Fresh sweep covers: ICLR 2026, NeurIPS 2025, ICML 2026 (main + Mech Interp Workshop), IEEE S&P 2026, USENIX Security 2026, arXiv cs.CL/cs.LG/cs.CR/cs.AI, Transformer Circuits Thread, LessWrong/Alignment Forum · Counts: 10 peer-reviewed (3 ICLR 2026 · 2 NeurIPS 2025 · 1 ICML 2026 Oral · 1 IEEE S&P 2026 · 1 USENIX Security 2026 · 2 ICML 2026 Workshop spotlights) · ~22 preprints · 1 lab blog
Window: 2026-07-11 to 2026-07-12 (preprints last 48h); peer-reviewed items not previously covered through July 2026 · Sources swept: EACL 2026 (ACL Anthology), AAAI 2026, ICML 2026 (main + Mech Interp Workshop), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), Semantic Scholar, Hugging Face Papers · Counts: 3 peer-reviewed · 7 preprints · 0 forum/blog
Window: 2026-07-10 to 2026-07-11; broader sweep for high-relevance preprints from May–July 2026 not covered in previous reports · Sources swept: ICML 2026 Mechanistic Interpretability Workshop (proceedings), OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), lab pages (NVIDIA Research, SLED/UMich), GitHub · Counts: 2 peer-reviewed (workshop) · 8 preprints · 0 forum/blog
Window: 2026-07-08 to 2026-07-10, plus high-relevance preprints not previously reported and newly-surfaced peer-reviewed work · Sources swept: Transformer Circuits Thread (Anthropic), USENIX Security 2026, OpenReview (ICML 2026), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Mechanistic Interpretability Workshop (today) · Counts: 1 peer-reviewed · 8 preprints · 1 forum/blog
Window: 2026-07-09, plus high-relevance preprints not previously reported and newly-surfaced peer-reviewed work · Sources swept: OpenReview (ICLR 2026), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), lab blogs · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: 2026-07-06 to 2026-07-08, plus newly-surfaced peer-reviewed work · Sources swept: OpenReview (ICML 2026, IEEE S&P 2026), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ACL Anthology, lab blogs · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: 2026-07-05 to 2026-07-07, plus peer-reviewed work not previously covered · Sources swept: OpenReview (ICLR 2026, NeurIPS 2025), arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ACL Anthology, lab blogs · Counts: 4 peer-reviewed · 6 preprints · 0 forum/blog
Window: June 30 – July 5, 2026 · Fresh sweep covers: ICLR 2026, ICML 2026 (main + Mech Interp Workshop + DL for Code Workshop), ECCV 2026, ACL 2026, arXiv cs.CL/cs.LG/cs.CR/cs.AI, LessWrong/Alignment Forum, lab blogs · Counts: 13 peer-reviewed (3 ICLR 2026, 4 ICML 2026 main, 3 ICML 2026 workshops, 2 ECCV 2026, 1 ACL 2026) · 24 preprints · 0 forum/blog posts new enough to rank
Window: July 3–5 2026 (arXiv 2607.xxxxx primary); extended to uncovered June–May 2026 preprints not surfaced in prior daily sweeps · Sources swept: ICML 2026 Mech Interp Workshop, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), LessWrong/Alignment Forum, Google DeepMind blog · Counts: 1 peer-reviewed · 9 preprints · 0 forum/blog
Window: July 2–4 2026 (arXiv 2607.xxxxx primary); extended to uncovered May–June 2026 preprints (2605–2606) due to reduced US-holiday submissions on July 4 · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Workshop on Mechanistic Interpretability, ECCV 2026, LessWrong/Alignment Forum · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: July 1–3 2026 (arXiv 2607.xxxxx primary); supplemented by June 2026 preprints (2606.xxxxx) not covered in prior dailies, plus one May 2026 dLLM-security paper first to reach the radar · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 Workshop on Mechanistic Interpretability, LessWrong/Alignment Forum · Counts: 2 peer-reviewed (ICML 2026 MI Workshop) · 8 preprints · 0 forum/blog
Window: ~June 4 – July 2, 2026 (preprints not covered in prior dailies); ICLR 2026 newly surfaced accepted papers · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), LessWrong/Alignment Forum, lab blogs · Counts: 2 peer-reviewed · 8 preprints · 0 forum/blog
Window: ~June 18 – July 1 2026 (preprints); ICLR 2026 & ICML 2026 newly presented peer-reviewed papers · Sources swept: OpenReview, arXiv (cs.CL/cs.LG/cs.CR/cs.AI), ICML 2026 virtual, ACL Anthology, ECCV 2026, Alignment Forum · Counts: 6 peer-reviewed · 4 preprints · 0 forum/blog
Compiled: 2026-06-30 · Window: posted/accepted Jul 2024 – Jun 2026 (seminal older anchors flagged) · Counts: ~27 items · peer-reviewed: SEDD (ICML'24 best paper), MDLM (NeurIPS'24), DiffuLLaMA (ICLR'25), Block Diffusion (ICLR'25 oral), Fast-dLLM (ICLR'26), DIJA/A2D/DiffuGuard (ICLR'26); rest preprints/industry reports.
Compiled: 2026-06-30 · Window: posted or accepted June 2025 – June 2026 · Coverage: ~110 unique items total — Part I ~52 (topic sweeps, §1–§8) + Part II ~58 (completeness additions, §9–§20, incl. the previously-missed AI Control area). ~40 peer-reviewed (incl. NeurIPS/ICML/ICLR orals & spotlights, ACL/EMNLP outstanding, S&P/USENIX distinguished); rest landmark preprints / major lab releases.