SynthesesWiki Lint Report — 2026-07-05DGX agentAutomated lint: 51 errors, 15 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-06-28DGX agentAutomated lint: 49 errors, 14 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-06-22DGX agentAutomated lint: 48 errors, 13 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-06-07DGX agentAutomated lint: 47 errors, 12 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-04-26DGX agentAutomated lint: 44 errors, 10 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-05-03DGX agentAutomated lint: 45 errors, 11 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-04-19DGX agentAutomated lint: 43 errors, 9 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-07-19DGX agentAutomated lint: 20 errors, 8743 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-07-15DGX agentAutomated lint: 26 errors, 6728 warnings, 3 info+Add to research
SynthesesWiki Lint Report — 2026-04-12DGX agentAutomated lint: 34 errors, 0 warnings, 3 info+Add to research
Model ReleasesMetaLint: Easy-to-Hard Generalization for Code LintingDGX agentarXiv:2507.11687v4 Announce Type: replace-cross Abstract: Large language models excel at code generation but struggle with code linting, particularly in generalizing to unseen or evolving best practic+Add to research
ResearchJust merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluen…DGX agentJust merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluency in Google's new open-source visual-identity format the mo+Add to research
Local Aiscicode-lint: Detecting Methodology Bugs in Scientific Python Code with LLM-Generated PatternsDGX agentarXiv:2603.17893v2 Announce Type: replace-cross Abstract: Methodology bugs in scientific Python code produce plausible but incorrect results that traditional linters and static analysis tools cannot d+Add to research
Model Releasessciwrite-lint: Verification Infrastructure for the Age of Science Vibe-WritingDGX agentarXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but+Add to research
Model ReleasesSomething I have been thinking about: in the past, the best engineers I knew spent a lot of time automating their work in various ways. Bett…DGX agentSomething I have been thinking about: in the past, the best engineers I knew spent a lot of time automating their work in various ways. Better vim/emacs automations, writing lint rules to catch repeat+Add to research
Model Releasesb10153DGX agentmodel: Add support for Nanbeige4.2 (#25994) support nanbeige4.2 model fix fix flake8 Lint check fix loop bound check and drop redundant head_dim Co-authored-by: root lizongqiang@kanzhun.com Website: h+Add to research
Model ReleasesRuff v0.16.0DGX agentRuff v0.16.0 Astral shipped a significant new version of their Ruff Python linting tool a few days ago on July 23rd. I noticed today because my various CI jobs all started failing thanks to new defaul+Add to research
ToolsNative Deployment Checks are now availableDGX agentNative Deployment Checks can now be run on every Vercel deployment, processing linting and typechecking in parallel with the build process. These built-in checks integrate with existing integrations (+Add to research
Local AiSemaDiff: Identifying Semantic-Changing Commits with Generated Code and TestsDGX agentarXiv:2607.13111v1 Announce Type: cross Abstract: Distinguishing semantic-preserving commits from changing ones remains an open challenge in software repository mining. While existing approaches detec+Add to research
AgentsQuality and Security Signals in AI-Generated Python Refactoring Pull RequestsDGX agentarXiv:2605.21453v1 Announce Type: cross Abstract: As AI agents increasingly contribute to code development and maintenance, there is still limited empirical evidence on the quality and risk characteri+Add to research
Model Releasesb10369DGX agentmtmd: support pocket-tts (#26871) adapt the api text model ok working impl, need verify and clean up mtmd: build the pocket-tts transposed convolutions as GEMM + col2im ggml_conv_transpose_1d has no g+Add to research
ResearchWeighted Sequential Bayesian Inference for Non-Stationary Linear Contextual BanditsDGX agentarXiv:2307.03587v4 Announce Type: replace Abstract: In non-stationary linear contextual bandits, existing efficient algorithms typically rely on the Weighted Regularized Least-Squares (WRLS) estimator+Add to research
SafetyWhat Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md FilesDGX agentarXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age+Add to research
Model Releasesb10270DGX agentmtmd: support Qwen3-TTS (note: breaking change to llama-tts binary) (#26254) convert text model main model load ok convert encoder ok speaker encoder loading ok speaker enc graph adapt vocab for backb+Add to research
Model ReleasesLFM2.5-Encoders: Fast at Long Context, Even on CPUDGX agentLFM2.5-Encoder is a family of multilingual bidirectional encoders built on the LFM2 architecture, available in two sizes: LFM2.5-Encoder-230M — a lightweight encoder for tight latency and memory budge+Add to research
Model Releasesb10142DGX agentmtmd: Add Vision Support for Minimax-M3 (#25113) Add preliminary MiniMax-M3 support Text-only port that re-uses existing components: MiniMax-M2 style GQA with per-head QK-norm and partial rotary, Deep+Add to research
Model ReleasesThe LangChain podcast where @hwchase17 interviews agent builders is full of alpha. Recent one with @EnoReyes was the best one yet. Sharing m…DGX agentThe LangChain podcast where @hwchase17 interviews agent builders is full of alpha. Recent one with @EnoReyes was the best one yet. Sharing my unstructured notes: Eno keeps bringing back some core conc+Add to research
ResearchA Preliminary Study on Explaining Risk of Code Changes using LLM-Based Prediction ModelsDGX agentarXiv:2607.02782v1 Announce Type: cross Abstract: Predictions by machine learning (ML) and artificial intelligence (AI) models are often received skeptically unless they are paired with intelligible e+Add to research
AgentsCheap Code, Costly Judgment: A Case Study on Governable Agentic Software EngineeringDGX agentarXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low+Add to research
Model ReleasesFrom Tool Connection to Execution Control: Benchmarking Security Invariants in MCP-Style Agent RuntimesDGX agentarXiv:2606.29073v1 Announce Type: cross Abstract: Model Context Protocol (MCP)-style ecosystems give language-model applications a practical connection layer for tools, resources, prompts, and transpo+Add to research
Model ReleasesBeyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMsDGX agentarXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va+Add to research
Model ReleasesImproving Small Language Models for Code Generation with Reinforcement Learning from Verification FeedbackDGX agentarXiv:2605.30478v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) trains language models using programmatically checkable signals such as unit-test outcomes, enab+Add to research
SafetyICAN-Deploy: Identity-Stable Canary Deployment for Safety-Critical Embodied AgentsDGX agentarXiv:2605.28097v1 Announce Type: new Abstract: Canary deployment routes a fraction of traffic to a new software version, monitors metrics, and rolls back on regression. Mainstream controllers (Argo R+Add to research
Model ReleasesShip code within minutes with the Gemini CLI DevOps ExtensionDGX agentWith AI coding tools like Antigravity and Claude Code, I can build a working web app in record time. But deploying it? That's where I'd historically lose the rest of the afternoon to Dockerfiles, IAM +Add to research
SafetyChipCraftBrain: Validation-First RTL Generation via Multi-Agent OrchestrationDGX agentarXiv:2604.19856v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise for generating Register-Transfer Level (RTL) code from natural language specifications, but single-shot gene+Add to research
SafetyArch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware DesignDGX agentarXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as+Add to research