githubReviewBench: An open benchmark for AI code reviewgithubdeveloper-newsAI & MLabout 4 hours ago▲ 56
githubCopilot code review: API support and new default effort levelgithubchangelogImprovement3 days ago▲ 56
TMtechmemeOpenAI announces new features for Codex, including reusable cloud development environments, a refreshed Codex CLI, a new code review experience, and more (Sarah Perez/TechCrunch)industry6 days ago▲ 24Read story→
rssOpenAI gives Codex reusable cloud environments that work across devicestechcrunchindustryAI6 days ago▲ 44Read story→
arXivarxivFinAutoRubric: Expert-Guided Automatic Rubric Generation for Evaluating Financial Research Agentsarxivcs.AIq-fin.CP7 days ago▲ 43Read story→
YhackernewsThere is more to code review than (automatable) detectionhackernews9 days ago▲ 15Read story→
YhackernewsShow HN: Foremerge – Catch Intent Conflicts Between Parallel Coding Agentshackernews14 days ago▲ 10Read story→
arXivarxivQuantifying Overclaiming Propensity in Frontier LLM Agentsarxivcs.SEcs.AI18 days ago▲ 43Read story→
YhackernewsGPT-5.6 Luna vs. GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review?hackernews21 days ago▲ 52Read story→
githubAuto-resolution and analysis updates in Copilot code reviewgithubchangelogImprovement24 days ago▲ 56
rssCognition helps Devin test its own work with GPT‑6 Astraopenaipublisher:openaitier:official24 days ago▲ 54Read story→
🦞lobstersIs it too much to ask devs to use AI to review their hand-crafted code?developer-newsprogrammingvibecodingabout 1 month ago▲ 14Read story→
YhackernewsGPT-6 Astra in code review: Gains, privacy, and costhackernewsabout 1 month ago▲ 4Read story→