AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlog
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
Safety

A true exponential!

DGX agent

A true exponential! Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. https://www.thelancet.com/journals/lancet/

safetygary-marcus--x
12 May 2026
Model Releases

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

model-releasesarxiv-cs-ai
12 May 2026
Safety

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

DGX agent

arXiv:2605.01289v1 Announce Type: new Abstract: This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuat

safetyarxiv-cs-ro
5 May 2026
Industry

Oh?

DGX agent

Oh? Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological cons

industryemad-mostaque--x
23 Apr 2026
Hardware

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind…

DGX agent

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind paper on test-time compute scaling. Here's what the agent d

hardwareclem-delangue--x
23 Apr 2026
Industry

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One ch…

DGX agent

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological constant

industryemad-mostaque--x
22 Apr 2026
Industry

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra o…

DGX agent

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra of relativity🌌 No new physics or math needed🧮 Possibly the mo

industryemad-mostaque--x
22 Apr 2026
Safety

Cooperative Coevolution versus Monolithic Evolutionary Search for Semi-Supervised Tabular Classification

DGX agent

arXiv:2604.16412v1 Announce Type: cross Abstract: This paper studies semi-supervised tabular classification in the extreme low-label regime using lightweight base learners. The paper proposes a cooper

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

DGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

model-releasesclem-delangue--x
21 Apr 2026
Research

[ECCV2026] Workshop notification of reject/accept[D]

DGX agent

This Reddit thread on r/MachineLearning discusses the workshop paper accept/reject notifications for ECCV 2026, the European Conference on Computer Vision (ECCV), a biennial premier research conferenc

researchr-machinelearning
13 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

DGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

safetyarxiv-cs-ro
10 Apr 2026
Research

[P] citracer: a small CLI tool to trace where a concept comes from in a citation graph

DGX agent

`citracer` is a small Python CLI tool available on PyPI that traces citation chains for any keyword across research papers, helping researchers identify where a concept originates in a citation gra...

researchr-machinelearning
8 Apr 2026
Research

MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems, Solutions, and Rationales

DGX agent

arXiv:2608.10974v1 Announce Type: new Abstract: Scientific papers contain fine-grained records of problem solving: authors mention technical obstacles and methods that were used to address them, often

researcharxiv-cs-cl
12 Aug 2026
Research

Tools to Explain Neural Networks for Power System Dynamics

DGX agent

arXiv:2608.08048v1 Announce Type: cross Abstract: This paper presents, for the first time in power systems literature to our knowledge, analytical tools to explain the training performance of machine

researcharxiv-cs-ai
11 Aug 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Safety

Challenges in Evaluating Explanation Methods for Static and Evolving Data

DGX agent

arXiv:2608.06351v1 Announce Type: new Abstract: This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through

safetyarxiv-cs-ai
7 Aug 2026
Research

Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model

DGX agent

arXiv:2608.03629v1 Announce Type: new Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an idealized model where a conditional computation is carried

researcharxiv-cs-ai
5 Aug 2026
Model Releases

How Closely Do LLM Reviews Align with Human Peer Review?

DGX agent

arXiv:2608.03659v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate scientific reviews, yet existing evaluations rarely examine whether different providers

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Applications

AI systems and the reproduction of (standard) language ideologies in World Englishes

DGX agent

arXiv:2607.28528v1 Announce Type: new Abstract: The rapid growth of large language models (LLMs) has resurrected age-old questions in sociolinguistics and world Englishes, such as who decides what cou

applicationsarxiv-cs-cl
31 Jul 2026
Model Releases

Nanbeige4.2-3B: I'm not impressed

DGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

model-releasesr-localllama
30 Jul 2026
Research

Measuring the State of Open Science in Transportation Using Large Language Models

DGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

researcharxiv-cs-ai
29 Jul 2026
Applications

Artificial Intelligence and Innovation Ecosystem: Evolutionary Developments, Challenges, and Future Directions

DGX agent

arXiv:2607.24589v1 Announce Type: new Abstract: The development of the Innovative Ecosystem (IE) presents a new paradigm for economic integration, collaborative advancement, and shared achievements. T

applicationsarxiv-cs-ai
28 Jul 2026
Model Releases

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate…

DGX agent

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate context every turn. The usual fix compresses on a token thr

model-releasesdair-ai--x
28 Jul 2026
Model Releases

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

DGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

model-releasesarxiv-cs-ai
28 Jul 2026
Local Ai

VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy

DGX agent

arXiv:2607.23006v1 Announce Type: cross Abstract: Scientific question answering requires a retrieval system to solve two distinct problems: identifying which papers are relevant and locating the suppo

local-aiarxiv-cs-ai
28 Jul 2026
Safety

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

DGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

safetyarxiv-cs-ai
16 Jul 2026
Safety

Introducing Human-Centeredness in AI-Assisted Lexicography

DGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

safetyarxiv-cs-ai
16 Jul 2026
Safety

AAAI-26 Dual Submissions: Novel Challenges

DGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

safetyarxiv-cs-ai
15 Jul 2026
Research

Metrics or Mirage? An Audit of Evaluation Inconsistencies in Colonoscopy Polyp Segmentation Benchmarks

DGX agent

arXiv:2607.08203v1 Announce Type: new Abstract: Progress in colonoscopy polyp segmentation is routinely reported through leaderboard comparisons on a small set of public benchmarks. We argue that this

researcharxiv-cs-cv
10 Jul 2026
Research

Evaluating RAG Metrics in Applied Contexts: An Experiment, Its Findings and Its Limitations

DGX agent

arXiv:2607.07302v1 Announce Type: new Abstract: This paper reports an empirical study evaluating the relevance of several RAG metrics. The experiment is based on a question-answering dataset created b

researcharxiv-cs-cl
9 Jul 2026
Model Releases

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of A…

DGX agent

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and scalin

model-releasesdair-ai--x
2 Jul 2026
Agents

Multi-Turn Agentic Scientific Literature Search via Workflow Induction

DGX agent

arXiv:2607.00597v1 Announce Type: new Abstract: Scientific literature search often requires more than retrieving papers from a single query: users' intents are underspecified, preference-dependent, an

agentsarxiv-cs-cl
2 Jul 2026
Model Releases

Rethinking the Role of Feature Engineering and Learning Strategies in Few-Shot Hidden Emotion Recognition

DGX agent

arXiv:2606.31249v1 Announce Type: new Abstract: In this paper, we present the solution developed by our team, XInsight Lab, which achieved first place in Track 3 of the 4th EI-MIGA-IJCAI Challenge wit

model-releasesarxiv-cs-cv
1 Jul 2026
Research

Exploring Motivations for Algorithm Mention in the Domain of Natural Language Processing: A Deep Learning Approach

DGX agent

arXiv:2606.29859v1 Announce Type: cross Abstract: With the rise of data-intensive science, algorithms have become central to scientific research. In academic papers, algorithms are mentioned for diffe

researcharxiv-cs-ai
30 Jun 2026
Tutorials

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways fo…

DGX agent

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways for AI ethicists and educators: The study analyzed data from 26

tutorialsgary-marcus--x
25 Jun 2026
Research

Improved Representation of Matrix Lie Group Operations through Tensor Notation

DGX agent

arXiv:2606.10289v1 Announce Type: new Abstract: Several recent papers have demonstrated the utility of using Lie groups within estimation problems, yielding improved accuracy and consistency. This pap

researcharxiv-cs-ro
10 Jun 2026
Model Releases

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

DGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Extending Ontologies: From Dense Embeddings to Hybrid Quantum-Fuzzy Systems

DGX agent

arXiv:2606.08658v1 Announce Type: new Abstract: LLMs have revolutionized knowledge representation and retrieval, but lack the explicit modeling that knowledge ontologies possess. This paper surveys th

researcharxiv-cs-ai
9 Jun 2026
Model Releases

The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition

DGX agent

arXiv:2606.00545v1 Announce Type: new Abstract: Post-trained language models can recognize their own outputs from a sentence or two out of context. In a companion paper itep{jack2026twomodes} we showe

model-releasesarxiv-cs-lg
2 Jun 2026
Applications

Advances and Challenges in Meta-Learning: A Technical Review

DGX agent

arXiv:2307.04722v2 Announce Type: replace Abstract: Meta-learning empowers learning systems with the ability to acquire knowledge from multiple tasks, enabling faster adaptation and generalization to

applicationsarxiv-cs-lg
1 Jun 2026
Safety

What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025

DGX agent

arXiv:2601.07648v2 Announce Type: replace Abstract: As Natural Language Generation (NLG) dominates modern NLP, scalable evaluation remains a critical bottleneck. Consequently, LLM-as-a-judge (LaaJ) ad

safetyarxiv-cs-cl
28 May 2026
Model Releases

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some int…

DGX agent

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some interesting tidbits. I summarized some of them below: 1. Full a

model-releasessebastian-raschka--x
27 May 2026
Agents

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

DGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

agentsarxiv-cs-cl
26 May 2026
Tutorials

Designing Singing Syllabi with Virtual Avatars: AI-Assisted Syllabus Reauthoring

DGX agent

arXiv:2508.11872v3 Announce Type: replace-cross Abstract: Traditional syllabi often function as static reference documents rather than engaging introductions to a course. In practical teaching, we obs

tutorialsarxiv-cs-ai
26 May 2026
Research

Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization

DGX agent

arXiv:2605.25203v1 Announce Type: cross Abstract: We apply the influence-adaptive Walsh geometry of a companion theory paper (arXiv:2605.01637) to extreme low-bit weight-only LLM quantization. The rec

researcharxiv-cs-ai
26 May 2026
Model Releases

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

DGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

model-releasesarxiv-cs-ai
26 May 2026
← Previous
1…89101112…252
Next →