AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
21,937 results
Safety

What We Know about Responsible AI Practices in Industry: A Half Decade of Empirical Research

DGX agent

arXiv:2608.10431v1 Announce Type: cross Abstract: Responsible AI (RAI) has become a central concern for technology companies, regulators, and the public. How industry practitioners interpret, implemen

safetyarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Artificial Intelligence Can Match Domain Experts in Evidence Extraction and Critical Appraisal of Microbial Oncogenesis Research Publications

DGX agent

arXiv:2608.07250v1 Announce Type: cross Abstract: Confirmed oncogenic microbes contribute significantly to cancer burden. Identifying novel microbial oncogenicity could yield strategies that will redu

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

DGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Algorithm Discovery

DGX agent

arXiv:2602.13769v3 Announce Type: replace Abstract: Automating heuristic design in complex, experiment-driven domains requires more than iterative mutation of solution algorithms. Current LLM-based ev

agentsarxiv-cs-ai
5 Aug 2026
Safety

Reviewer Scores Are Not Comparable Across Research Areas in ML Peer Review

DGX agent

arXiv:2607.27209v1 Announce Type: cross Abstract: Peer review at ML conferences increasingly relies on reviewer scores as the primary decision instrument. As submissions have scaled from thousands to

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation

DGX agent

arXiv:2607.25881v1 Announce Type: new Abstract: We investigate how well large language models (LLMs) can assist scientific project planning and proposal evaluation. One-page project plans were indepen

model-releasesarxiv-cs-cl
29 Jul 2026
Research

Measuring the State of Open Science in Transportation Using Large Language Models

DGX agent

arXiv:2601.14429v2 Announce Type: replace-cross Abstract: Open science initiatives have strengthened scientific integrity and accelerated research progress across many fields, but the state of their p

researcharxiv-cs-ai
29 Jul 2026
Applications

'We'll have to see how it works': An interview study to understand collaborative practices in interdisciplinary artificial intelligence and healthcare research

DGX agent

arXiv:2311.18424v3 Announce Type: replace-cross Abstract: Developing artificial intelligence (AI) algorithms for healthcare is a collaborative effort, bringing data scientists, clinicians, patients an

applicationsarxiv-cs-ai
29 Jul 2026
Research

Knowledge Graph and Accurate Portrait Construction of Scientific and Technological Academic Conferences

DGX agent

arXiv:2204.04888v2 Announce Type: replace-cross Abstract: In recent years, with the continuous progress of science and technology, the number of scientific research achievements has increased rapidly.

researcharxiv-cs-ai
10 Jul 2026
Agents

Beyond the Library: An Agentic Framework for Autoformalizing Research Mathematics

DGX agent

arXiv:2606.31134v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated exceptional capabilities in mathematical reasoning, they frequently produce subtle errors that evad

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

Can LLMs Prove Robotic Path Planning Optimality? A Benchmark for Research-Level Algorithm Verification

DGX agent

arXiv:2603.19464v2 Announce Type: replace Abstract: Robotic path planning problems are often NP-hard, and practical solutions typically rely on approximation algorithms with provable performance guara

model-releasesarxiv-cs-ro
30 Jun 2026
Model Releases

Research Entity Extraction and Topic Detection from UKRI Grant Proposals

DGX agent

arXiv:2606.30304v1 Announce Type: cross Abstract: This paper presents preliminary findings from a UKRI-funded Metascience project comparing three LLM-based approaches, GPT-4o, Mistral, and a bespoke a

model-releasesarxiv-cs-ai
30 Jun 2026
Applications

Large-scale semantic mapping of learner agency and autonomy reveals what measurement and generative AI research overlook

DGX agent

arXiv:2606.10881v1 Announce Type: new Abstract: Learner agency and autonomy are foundational to personal development, yet a pervasive 'jingle-jangle' fallacy (i.e. identical terms denoting different c

applicationsarxiv-cs-ai
10 Jun 2026
Agents

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

DGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

agentsarxiv-cs-ai
9 Jun 2026
Applications

Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariates

DGX agent

arXiv:2606.03029v1 Announce Type: cross Abstract: A core goal of computational social science is to discover interpretable differences in how language varies across outcomes of interest, such as polit

applicationsarxiv-cs-ai
3 Jun 2026
Model Releases

GTBench: A Curriculum-Grounded Benchmark for Evaluating LLMs as Mathematical Research Assistants in Graph Theory

DGX agent

arXiv:2606.03144v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as self-study assistants in technical disciplines, yet their reliability as mathematical reasoning as

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Bridging Evolutionary Algorithms and Reinforcement Learning: A Comprehensive Survey on Hybrid Algorithms

DGX agent

arXiv:2401.11963v5 Announce Type: replace-cross Abstract: Evolutionary Reinforcement Learning (ERL), which integrates Evolutionary Algorithms (EAs) and Reinforcement Learning (RL) for optimization, ha

researcharxiv-cs-ai
26 May 2026
Research

SciNet: Evaluating AI Agents in Relation-Aware Scientific Literature Retrieval

DGX agent

arXiv:2601.03260v2 Announce Type: replace-cross Abstract: AI agents have seen widespread adoption in information retrieval for scientific research, giving rise to tools such as Deep Research. However,

researcharxiv-cs-cl
25 May 2026
Safety

Assured autonomy: How operations research powers and orchestrates generative AI systems

DGX agent

arXiv:2512.23978v2 Announce Type: replace Abstract: Generative artificial intelligence (GenAI) is shifting from conversational assistants toward agentic systems -- autonomous decision-making systems t

safetyarxiv-cs-lg
19 May 2026
Model Releases

BoLT: A Benchmark to Democratize Black-box Optimization Research for Expensive LLM Tasks

DGX agent

arXiv:2605.17000v1 Announce Type: cross Abstract: Optimization of LLM training and inference configurations, such as hyperparameters, data mixtures, and prompts, is critical to performance, but it is

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Accelerating battery research with an AI interface between FINALES and Kadi4Mat

DGX agent

arXiv:2605.00909v1 Announce Type: cross Abstract: The time-consuming formation process critically impacts the longevity of sodium-ion coin cells and End Of Life (EOL) performance. This study aims to o

model-releasesarxiv-cs-lg
5 May 2026
Agents

Bolzano: Case Studies in LLM-Assisted Mathematical Research

DGX agent

arXiv:2604.16989v1 Announce Type: new Abstract: We report new results on six problems in mathematics and theoretical computer science, produced with the assistance of Bolzano, an open-source multi-age

agentsarxiv-cs-cl
21 Apr 2026
Agents

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

DGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

agentsarxiv-cs-ro
16 Apr 2026
Model Releases

Competing with AI Scientists: Agent-Driven Approach to Astrophysics Research

DGX agent

arXiv:2604.09621v1 Announce Type: new Abstract: We present an agent-driven approach to the construction of parameter inference pipelines for scientific data analysis. Our method leverages a multi-agen

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Eleven Years of BRACIS: A Meta-Scientific Study of the Brazilian Conference on Intelligent Systems

DGX agent

arXiv:2608.09964v1 Announce Type: cross Abstract: The Brazilian Conference on Intelligent Systems (BRACIS) is the main national venue for Artificial Intelligence research in Brazil, hosted by the Braz

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Towards Researcher Agents for Knowledge-Graph Question Answering

DGX agent

arXiv:2608.07700v1 Announce Type: new Abstract: Translating a natural-language question into a SPARQL query that can be executed against a large knowledge graph requires resolving lexical ambiguity, g

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research

DGX agent

arXiv:2605.12153v2 Announce Type: replace-cross Abstract: We present the Curated Industrial Developer Repository (CIDR), a large-scale dataset of real-world software repositories collected from indust

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

What Could the Agent See at 19:05? Generating Temporal Enterprise Scenarios from Real Research and Replaying Them to Evaluate Agents

DGX agent

arXiv:2608.01042v1 Announce Type: cross Abstract: Enterprise AI agents act across many apps whose data changes continuously, so an answer is correct only relative to what data existed and who could se

agentsarxiv-cs-lg
4 Aug 2026
Safety

Large Language Model for Operations Research Formulation Selection in Multi-Warehouse Inventory Allocation

DGX agent

arXiv:2607.25956v1 Announce Type: new Abstract: Multi-warehouse inventory allocation is typically formulated as a mixed-integer programming (MIP) problem, yet no single formulation consistently matche

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

Bringing Back Rule Induction to Fluid Intelligence Research? An Initial Validation of the ARC-AGI Benchmark in Humans

DGX agent

arXiv:2607.11263v2 Announce Type: replace Abstract: Two competing perspectives on fluid intelligence (gf) measures propose that performance is primarily constrained either by working memory capacity o

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Do You Need a Frontier Model as a Citation Verifier? Benchmarking Rubric LLMs for Deep-Research Source Attribution

DGX agent

arXiv:2607.08700v1 Announce Type: new Abstract: Reinforcement learning increasingly relies on an LLM judge to score each rubric criterion, and that judge acts as the reward model during training. Befo

model-releasesarxiv-cs-cl
10 Jul 2026
Model Releases

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

DGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

DGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Unveiling Novelty Evolution in the field of Library and Information Science in China

DGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

researcharxiv-cs-cl
30 Jun 2026
Agents

Agentic Software Engineering: Foundational Pillars and a Research Roadmap

DGX agent

arXiv:2509.06216v3 Announce Type: replace-cross Abstract: Agentic Software Engineering (SE 3.0) represents a new era where intelligent agents are tasked not with simple code generation, but with achie

agentsarxiv-cs-ai
25 Jun 2026
Agents

Democratizing and accelerating AI-driven pathology research through agentic intelligence

DGX agent

arXiv:2606.20677v1 Announce Type: cross Abstract: Computational pathology has advanced rapidly with the emergence of foundation models, yet widespread adoption remains limited by substantial technical

agentsarxiv-cs-cv
23 Jun 2026
Safety

Urban Heat MiniCubes: An AI-Ready dataset for urban heat research

DGX agent

arXiv:2606.11534v1 Announce Type: cross Abstract: Urban heat is amplified by impermeable surfaces and heterogeneous built environments, yet street-level variability remains difficult to quantify becau

safetyarxiv-cs-lg
11 Jun 2026
Model Releases

Evaluating Research-Level Math Proofs via Strict Step-Level Verification

DGX agent

arXiv:2606.10799v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle to rigorously verify complex mathematical proofs. Standard global evaluation approaches suffer from 'context poiso

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

What Fits (Into Few Tokens) Doesn't Overfit: Compression and Generalization in ML Research Agents

DGX agent

arXiv:2606.11045v1 Announce Type: new Abstract: Reusing a held-out benchmark adaptively should, in principle, invite overfitting. Yet benchmark-driven machine learning (ML) has produced surprisingly l

model-releasesarxiv-cs-ai
10 Jun 2026
Local Ai

Position: Genomic Model Research Must Move Beyond Anecdotal Evaluation of Interpretability Methods

DGX agent

arXiv:2606.07607v1 Announce Type: new Abstract: Advances in machine learning and computational power have unlocked the predictive potential of the human genome, yet biologists now demand that these mo

local-aiarxiv-cs-lg
9 Jun 2026
Applications

A machine-learning-assisted progressive digit-randomness screening framework for detecting non-random patterns in raw numerical research data

DGX agent

arXiv:2606.07128v1 Announce Type: new Abstract: Raw numerical datasets remain less systematically examined in integrity screening than images, plagiarism, or summary-statistic inconsistencies. We deve

applicationsarxiv-cs-lg
8 Jun 2026
Model Releases

CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions

DGX agent

arXiv:2606.06526v1 Announce Type: new Abstract: Large language models have made substantial progress on mathematical reasoning, but existing benchmarks typically evaluate well-specified problems with

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

AutoLab: Can Frontier Models Solve Long-Horizon Auto Research and Engineering Tasks?

DGX agent

arXiv:2606.05080v1 Announce Type: new Abstract: Scientific and engineering progress is fundamentally a long-horizon iterative process: proposing changes, running experiments, measuring outcomes, and c

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

DGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Rethinking Literature Search Evaluation: Deep Research Helps, and Human Citation Lists Are Not a Ground Truth

DGX agent

arXiv:2605.29234v1 Announce Type: new Abstract: We study large-scale literature search from two complementary angles: improving the retrieval pipeline, and stress-testing the human reference list as a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

ConvMemory: A Lightweight Learned Memory Reranker, a Negative Attribution Result, and a Research-Preview Conflict Editor

DGX agent

arXiv:2605.28062v1 Announce Type: new Abstract: We describe ConvMemory, a small 3.6M-parameter learned reranker for conversational long-term memory retrieval, trained with cross-encoder teacher superv

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

E3: Issue-Level Backtesting for Automated Research Critique

DGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…678910…458
Next →