AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
3 Jul 2026

PreScience: A Dataset and Benchmark for Scientific Forecasting

Model ReleasesDGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

30 Jun 2026

HyBIRD: Hyperbolic Bridge Retrieval and Diagnosis for Methodology Inspiration Retrieval

Model ReleasesDGX agent

arXiv:2606.28336v1 Announce Type: cross Abstract: Methodology Inspiration Retrieval (MIR) asks a system to retrieve prior papers whose methods can inspire a new research proposal. Unlike general scien

Exploring Motivations for Algorithm Mention in the Domain of Natural Language Processing: A Deep Learning Approach

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2606.29859v1 Announce Type: cross Abstract: With the rise of data-intensive science, algorithms have become central to scientific research. In academic papers, algorithms are mentioned for diffe

26 Jun 2026

The Governance Inversion Hypothesis: Why More AI Regulation May Produce Less Organisational Control

ResearchDGX agent

arXiv:2606.26117v1 Announce Type: cross Abstract: This paper introduces the Governance Inversion Hypothesis (GIH) to explain a growing paradox in artificial intelligence (AI) governance: under conditi

25 Jun 2026

ReviewGuard: Aligning LLM-Assisted Peer Review with Long-Term Scientific Impact

ResearchDGX agent

arXiv:2606.24892v1 Announce Type: cross Abstract: Peer review is central to scientific quality control, yet it can undervalue papers that later achieve substantial citation impact. While frontier larg

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways fo…

TutorialsDGX agent

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways for AI ethicists and educators: The study analyzed data from 26

10 Jun 2026

my weekend hobby: self improvement research

AgentsDGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

Improved Representation of Matrix Lie Group Operations through Tensor Notation

ResearchDGX agent

arXiv:2606.10289v1 Announce Type: new Abstract: Several recent papers have demonstrated the utility of using Lie groups within estimation problems, yielding improved accuracy and consistency. This pap

3 Jun 2026

Merit or networks? What decides where research is published

ApplicationsDGX agent

arXiv:2606.03763v1 Announce Type: cross Abstract: Does scientific publishing reward the quality of ideas or the advantage of connections? The question is universal to prestige-driven science, yet it h

29 May 2026

Review Arcade: On the Human Alignment and Gameability of LLM Reviews

SafetyDGX agent

arXiv:2605.28897v1 Announce Type: new Abstract: LLM-generated reviews for scientific papers are gaining considerable traction and are even being officially piloted by major conferences. We have to ass

28 May 2026

AI Research Agents Narrow Scientific Exploration

AgentsDGX agent

arXiv:2605.27905v1 Announce Type: new Abstract: AI research agents can now generate research ideas, design experiments, run code, and draft papers, raising the possibility of large-scale AI-assisted s

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture'

AgentsDGX agent

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture' babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is t

What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025

SafetyDGX agent

arXiv:2601.07648v2 Announce Type: replace Abstract: As Natural Language Generation (NLG) dominates modern NLP, scalable evaluation remains a critical bottleneck. Consequently, LLM-as-a-judge (LaaJ) ad

25 May 2026

Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good

SafetyDGX agent

arXiv:2605.22995v1 Announce Type: cross Abstract: Agentic AI systems are increasingly proposed for social-good domains, often invoking the United Nations Sustainable Development Goals (SDGs) as a voca

22 May 2026

recommended reading.

AgentsDGX agent

recommended reading. babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems

14 May 2026

No One Knows the State of the Art in Geospatial Foundation Models

Model ReleasesDGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

12 May 2026

A true exponential!

SafetyDGX agent

A true exponential! Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. https://www.thelancet.com/journals/lancet/

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

Model ReleasesDGX agent

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

5 May 2026

Bi-Level Reinforcement Learning Control for an Underactuated Blimp via Center-of-Mass Reconfiguration

SafetyDGX agent

arXiv:2605.01289v1 Announce Type: new Abstract: This paper investigates goal-directed tracking control of underactuated blimps with center-of-mass (CoM) reconfiguration. Unlike conventional overactuat

23 Apr 2026

Oh?

IndustryDGX agent

Oh? Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological cons

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind…

HardwareDGX agent

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind paper on test-time compute scaling. Here's what the agent d

22 Apr 2026

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One ch…

IndustryDGX agent

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological constant

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra o…

IndustryDGX agent

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra of relativity🌌 No new physics or math needed🧮 Possibly the mo

21 Apr 2026

Cooperative Coevolution versus Monolithic Evolutionary Search for Semi-Supervised Tabular Classification

SafetyDGX agent

arXiv:2604.16412v1 Announce Type: cross Abstract: This paper studies semi-supervised tabular classification in the extreme low-label regime using lightweight base learners. The paper proposes a cooper

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

Model ReleasesDGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

13 Apr 2026

[ECCV2026] Workshop notification of reject/accept[D]

ResearchDGX agent

This Reddit thread on r/MachineLearning discusses the workshop paper accept/reject notifications for ECCV 2026, the European Conference on Computer Vision (ECCV), a biennial premier research conferenc

10 Apr 2026

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

Model ReleasesDGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

SafetyDGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

8 Apr 2026

[P] citracer: a small CLI tool to trace where a concept comes from in a citation graph

ResearchDGX agent

`citracer` is a small Python CLI tool available on PyPI that traces citation chains for any keyword across research papers, helping researchers identify where a concept originates in a citation gra...

12 Aug 2026

MUSE: A Full-Text Cross-Domain Knowledge Base of Scientific Problems, Solutions, and Rationales

ResearchDGX agent

arXiv:2608.10974v1 Announce Type: new Abstract: Scientific papers contain fine-grained records of problem solving: authors mention technical obstacles and methods that were used to address them, often

11 Aug 2026

Tools to Explain Neural Networks for Power System Dynamics

ResearchDGX agent

arXiv:2608.08048v1 Announce Type: cross Abstract: This paper presents, for the first time in power systems literature to our knowledge, analytical tools to explain the training performance of machine

10 Aug 2026

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

Model ReleasesDGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

7 Aug 2026

Challenges in Evaluating Explanation Methods for Static and Evolving Data

SafetyDGX agent

arXiv:2608.06351v1 Announce Type: new Abstract: This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through

5 Aug 2026

Cross-Layer Interaction under Weight-Space Ablation: A Closed-Form Attention Jacobian Bound and a Test on a Real Pretrained Model

ResearchDGX agent

arXiv:2608.03629v1 Announce Type: new Abstract: A companion paper studies when activation patching and weight-space ablation agree, inside an idealized model where a conditional computation is carried

How Closely Do LLM Reviews Align with Human Peer Review?

Model ReleasesDGX agent

arXiv:2608.03659v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate scientific reviews, yet existing evaluations rarely examine whether different providers

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

Model ReleasesDGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

31 Jul 2026

AI systems and the reproduction of (standard) language ideologies in World Englishes

ApplicationsDGX agent

arXiv:2607.28528v1 Announce Type: new Abstract: The rapid growth of large language models (LLMs) has resurrected age-old questions in sociolinguistics and world Englishes, such as who decides what cou

30 Jul 2026

Nanbeige4.2-3B: I'm not impressed

Model ReleasesDGX agent

I've tested Nanbeige-4.2-3B. On paper, the benchmarks promise it blows away Qwen3.5-9B and Gemma4-12B. My goal was to have something very light and fast to replace Qwen3.6-35B (or finetunes thereof) f

29 Jul 2026

Measuring the State of Open Science in Transportation Using Large Language Models

ResearchDGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

28 Jul 2026

Artificial Intelligence and Innovation Ecosystem: Evolutionary Developments, Challenges, and Future Directions

ApplicationsDGX agent

arXiv:2607.24589v1 Announce Type: new Abstract: The development of the Innovative Ecosystem (IE) presents a new paradigm for economic integration, collaborative advancement, and shared achievements. T

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate…

Model ReleasesDGX agent

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate context every turn. The usual fix compresses on a token thr

The Half-Lives of Generative-AI Evidence: A 40-Record Audit, a Claim-Currency Framework, and a Reflexive Case of Frontier-Model-Assisted Research

Model ReleasesDGX agent

arXiv:2607.24032v1 Announce Type: new Abstract: Generative-AI evaluations can become historical before publication, yet calendar age does not affect every conclusion equally. This paper has two linked

VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy

Local AiDGX agent

arXiv:2607.23006v1 Announce Type: cross Abstract: Scientific question answering requires a retrieval system to solve two distinct problems: identifying which papers are relevant and locating the suppo

16 Jul 2026

Final Authority in AI Governance: Frontier-Provider Sovereignty and Action-Centered Deployer Governance

SafetyDGX agent

arXiv:2607.13040v1 Announce Type: cross Abstract: This paper examines where final authority should sit once capable AI systems are embedded in organizational workflows. It compares two governance mode

Introducing Human-Centeredness in AI-Assisted Lexicography

SafetyDGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

15 Jul 2026

AAAI-26 Dual Submissions: Novel Challenges

SafetyDGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

10 Jul 2026

Metrics or Mirage? An Audit of Evaluation Inconsistencies in Colonoscopy Polyp Segmentation Benchmarks

ResearchDGX agent

arXiv:2607.08203v1 Announce Type: new Abstract: Progress in colonoscopy polyp segmentation is routinely reported through leaderboard comparisons on a small set of public benchmarks. We argue that this

9 Jul 2026

Evaluating RAG Metrics in Applied Contexts: An Experiment, Its Findings and Its Limitations

ResearchDGX agent

arXiv:2607.07302v1 Announce Type: new Abstract: This paper reports an empirical study evaluating the relevance of several RAG metrics. The experiment is based on a question-answering dataset created b

2 Jul 2026

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of A…

Model ReleasesDGX agent

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and scalin

Multi-Turn Agentic Scientific Literature Search via Workflow Induction

AgentsDGX agent

arXiv:2607.00597v1 Announce Type: new Abstract: Scientific literature search often requires more than retrieving papers from a single query: users' intents are underspecified, preference-dependent, an

1 Jul 2026

Rethinking the Role of Feature Engineering and Learning Strategies in Few-Shot Hidden Emotion Recognition

Model ReleasesDGX agent

arXiv:2606.31249v1 Announce Type: new Abstract: In this paper, we present the solution developed by our team, XInsight Lab, which achieved first place in Track 3 of the 4th EI-MIGA-IJCAI Challenge wit

9 Jun 2026

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

Model ReleasesDGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

Extending Ontologies: From Dense Embeddings to Hybrid Quantum-Fuzzy Systems

ResearchDGX agent

arXiv:2606.08658v1 Announce Type: new Abstract: LLMs have revolutionized knowledge representation and retrieval, but lack the explicit modeling that knowledge ontologies possess. This paper surveys th

2 Jun 2026

The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition

Model ReleasesDGX agent

arXiv:2606.00545v1 Announce Type: new Abstract: Post-trained language models can recognize their own outputs from a sentence or two out of context. In a companion paper itep{jack2026twomodes} we showe

1 Jun 2026

Advances and Challenges in Meta-Learning: A Technical Review

ApplicationsDGX agent

arXiv:2307.04722v2 Announce Type: replace Abstract: Meta-learning empowers learning systems with the ability to acquire knowledge from multiple tasks, enabling faster adaptation and generalization to

27 May 2026

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some int…

Model ReleasesDGX agent

The MiniMax M2 series was one of the most widely used open-weight LLM series earlier this year. Now, we got a technical report with some interesting tidbits. I summarized some of them below: 1. Full a

26 May 2026

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

AgentsDGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

Designing Singing Syllabi with Virtual Avatars: AI-Assisted Syllabus Reauthoring

TutorialsDGX agent

arXiv:2508.11872v3 Announce Type: replace-cross Abstract: Traditional syllabi often function as static reference documents rather than engaging introductions to a course. In practical teaching, we obs

Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization

ResearchDGX agent

arXiv:2605.25203v1 Announce Type: cross Abstract: We apply the influence-adaptive Walsh geometry of a companion theory paper (arXiv:2605.01637) to extreme low-bit weight-only LLM quantization. The rec

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

Model ReleasesDGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

← Previous
1…678910…201
Next →