AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,646 results
9 Jun 2026

Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting

Model ReleasesDGX agent

arXiv:2606.09809v1 Announce Type: new Abstract: AI evaluation results are produced at scale but reported inconsistently across leaderboards, model cards, benchmark papers, and company blogs. The cost

MIRAGE: Metadata-Integrated Repository Analysis and Guided Enhancement for MSR Datasets

SafetyDGX agent

arXiv:2606.07611v1 Announce Type: cross Abstract: This paper proposes an improved approach to the analysis of Mining Software Repositories (MSR) datasets via metadata enrichment, FAIRness assessment,

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

Model ReleasesDGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great…

Model ReleasesDGX agent

This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great and it's SOTA on everything by a margin but I'll add that *

this is the biggest wake-up call to protect and nourish open source AI if you don't build out sovereign and independent models+infra closed …

IndustryDGX agent

this is the biggest wake-up call to protect and nourish open source AI if you don't build out sovereign and independent models+infra closed labs will patronize you to an insulting degree mythos will b

Towards Personalized Bangla Book Recommendation: A Large-Scale Heterogeneous Book Graph Dataset

Model ReleasesDGX agent

arXiv:2602.12129v2 Announce Type: replace-cross Abstract: Personalized book recommendation in Bangla literature has been constrained by the lack of structured, large-scale, and publicly available data

Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing

AgentsDGX agent

arXiv:2606.08256v1 Announce Type: new Abstract: Verifiability, attribution, and reproducibility are foundational requirements of scientific knowledge, yet current publishing infrastructure does not en

8 Jun 2026

Are you sure? A Comprehensive and Comprehensible Survey of Uncertainty Quantification in Symbolic Regression

ApplicationsDGX agent

arXiv:2606.06567v1 Announce Type: new Abstract: Symbolic regression (SR) is a class of methods that systematically explore the space of mathematical functions to discover models that accurately captur

Measuring Agents in Production

AgentsDGX agent

arXiv:2512.04123v4 Announce Type: replace-cross Abstract: LLM-based agents already operate in production across many industries, yet we lack an understanding of what technical methods make deployments

Twelve quick tips for designing AI-driven HPC workflows

TutorialsDGX agent

arXiv:2606.07491v1 Announce Type: cross Abstract: High-performance computing (HPC) clusters remain the backbone of large-scale scientific computation, traditionally executing deterministic, linear pip

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://act…

AgentsDGX agent

two things ready to share from this weekend: 📖 http://learn.activegraph.ai interactive site teaching activegraph concepts blog: https://activegraph.ai/blog/introducing-learn 💻 AG coder (open source) r

7 Jun 2026

Russell Vought’s proposal to make politics, not peer review, the standard for NIH science grants would be a cataclysm for the American scien…

ApplicationsDGX agent

Russell Vought proposed replacing peer review with political criteria for NIH science grants, a change that would fundamentally compromise the scientific merit-based evaluation process. The proposal w

6 Jun 2026

Assessing the Geographic Diversity of AI's Platial Representations in Image Generation

SafetyDGX agent

arXiv:2606.05188v1 Announce Type: cross Abstract: (Gen)AI diversity is not merely an ethical issue. From the perspective of geographic information science (GIScience), it could be interpreted as a fun

Geographic Bias and Diversity in AI Evaluation

Model ReleasesDGX agent

arXiv:2606.05187v1 Announce Type: cross Abstract: Among the many challenges hindering the responsible development and deployment of AI, arguably none has faced more intense scrutiny than bias in its v

Scientists ejected from diabetes conference for distributing journal reprints

IndustryDGX agent

Five scientists, including the editor-in-chief of America's leading diabetes journal, were removed from the American Diabetes Association's annual conference in New Orleans on June 5, 2026, after dist

Zero knowledge verification for frontier AI training is possible

SafetyDGX agent

arXiv:2606.05433v1 Announce Type: new Abstract: Frontier AI governance frameworks increasingly use cumulative training compute as the primary criterion for designating high-impact models, but enforcem

5 Jun 2026

Benchmarking Open-Source Layout Detection Models for Data Snapshot Extraction from Institutional Documents

Model ReleasesDGX agent

arXiv:2606.06242v1 Announce Type: new Abstract: Institutional documents contain substantial amounts of operational and analytical information embedded within figures and tables. Current approaches for

Political appointees vetting science funding? Trump is doing to science what the Right accused previous administrations of doing. But in pre…

SafetyDGX agent

Political appointees vetting science funding? Trump is doing to science what the Right accused previous administrations of doing. But in previous administrations, there was no political control of res

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I…

ApplicationsDGX agent

Today, we are officially launching the Sakana AI RSI Lab in Tokyo to build open-ended, adaptive AI systems that collectively self-improve. I am incredibly proud of our team’s work over the past 2 year

4 Jun 2026

Can Crowdsourcing Survive the LLM Era? A Community Survey on Human Data Collection

TutorialsDGX agent

arXiv:2606.04924v1 Announce Type: new Abstract: The widespread use of Large Language Models (LLMs) as writing tools challenges the validity of crowdsourced data, as crowdworkers may outsource tasks to

Can Generalist Agents Automate Data Curation?

Model ReleasesDGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

DetectZoo: A Unified Toolkit for AI-Generated Content Detection Across Text, Audio, and Image Modalities

Model ReleasesDGX agent

arXiv:2606.04205v1 Announce Type: cross Abstract: The growing popularity and capacity of generative models have eroded the distinction between human and machine-generated content, motivating a growing

How do machines learn? Evaluating the AIcon2abs method

TutorialsDGX agent

arXiv:2401.07386v5 Announce Type: cross Abstract: This study expands on previous work that introduced the AIcon2abs method (AI from Concrete to Abstract: Demystifying Artificial Intelligence to the ge

Neetyabhas: A Framework for Uncertainty-Aware Public Policy Optimization in Rational Agent-Based Models

SafetyDGX agent

arXiv:2606.04562v1 Announce Type: new Abstract: Purpose The WHO's COVID-19 non-pharmaceutical interventions (e.g., lockdowns, vaccinations) effectively curb transmission but impose heavy economic stra

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

Model ReleasesDGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

UniCAD: A Unified Benchmark and Universal Model for Multi-Modal Multi-Task CAD

Model ReleasesDGX agent

arXiv:2606.05058v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) underpins modern engineering and manufacturing by enabling the creation of precise, editable 3D models. However, CAD resea

3 Jun 2026

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions

AgentsDGX agent

arXiv:2606.02859v1 Announce Type: cross Abstract: How can a population of agents self-orchestrate and self-adapt into stronger collective intelligence without centralized control? Inspired by Friedric

Explainable Forecasting of Scientific Breakthroughs from Concept Network Dynamics

SafetyDGX agent

arXiv:2606.03864v1 Announce Type: cross Abstract: We introduce an explainable machine-learning approach that forecasts the structural precursors of scientific breakthroughs -- the emergence and intens

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

Model ReleasesDGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

TriEval: A Resource-Efficient Pipeline for LLM Bias, Toxicity, and Truthfulness Assessment

Model ReleasesDGX agent

arXiv:2606.03036v1 Announce Type: new Abstract: LLMs have evolved from basic chatbots to the backbone of the AI ecosystem, now widely used in healthcare, schools, and government services. The domain-w

2 Jun 2026

A multimodal dataset of photoplethysmography and continuous behavioral responses to ASMR and nature videos

AgentsDGX agent

arXiv:2606.00752v1 Announce Type: new Abstract: Autonomous Sensory Meridian Response (ASMR) is a somatosensory phenomenon characterized by pleasant tingling sensations and cardiovascular slowing. Howe

Embedding-Space Diffusion for Zero-Shot Environmental Sound Classification

Model ReleasesDGX agent

arXiv:2412.03771v3 Announce Type: replace-cross Abstract: Zero-shot learning enables models to generalise to unseen classes by leveraging semantic information, bridging the gap between training and te

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

Model ReleasesDGX agent

arXiv:2606.00750v1 Announce Type: new Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing

MARFT: Multi-Agent Reinforcement Fine-Tuning

AgentsDGX agent

arXiv:2504.16129v5 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (LaMAS) have demonstrated strong capabilities on complex agentic tasks requiring multifac

MMDG-Bench: A Benchmark for Multimodal Domain Generalization

Model ReleasesDGX agent

arXiv:2606.00891v1 Announce Type: new Abstract: Multi-modal Domain Generalization (MMDG) seeks to leverage complementary modalities to enhance model robustness on unseen domains. Despite extensive pro

PaperVoyager : Building Interactive Web with Visual Language Models

Model ReleasesDGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

SafetyDGX agent

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

Property Prediction of Stacked Bilayer Materials: A Multimodal Learning Approach

ApplicationsDGX agent

arXiv:2606.01012v1 Announce Type: new Abstract: AI for materials science is a critical topic within AI for science, aiming to accelerate materials discovery and produce accurate property predictions.

RenoBench: A Citation Parsing Benchmark

Model ReleasesDGX agent

arXiv:2603.25640v2 Announce Type: replace-cross Abstract: Accurate parsing of citations is necessary for machine-readable scholarly infrastructure. But, despite sustained interest in this problem, exi

1 Jun 2026

Advances and Challenges in Meta-Learning: A Technical Review

ApplicationsDGX agent

arXiv:2307.04722v2 Announce Type: replace Abstract: Meta-learning empowers learning systems with the ability to acquire knowledge from multiple tasks, enabling faster adaptation and generalization to

Extending the UXR Point of View Pyramid: A Generative AI-Augmented Methodology for Human-Centred AI Systems

SafetyDGX agent

arXiv:2605.31143v1 Announce Type: cross Abstract: Rising household debt and cost-of-living pressures in the United Kingdom have intensified the role of AI-driven financial technologies in mediating cr

GEM-Bench: A Benchmark for Ad-Injected Response Generation within Generative Engine Marketing

Model ReleasesDGX agent

arXiv:2509.14221v3 Announce Type: replace-cross Abstract: Generative Engine Marketing (GEM) is an emerging ecosystem for monetizing generative engines, such as LLM-based chatbots, by seamlessly integr

Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

AgentsDGX agent

arXiv:2605.30677v1 Announce Type: cross Abstract: Agentic software reverse engineering systems are vulnerable to prompt injection attacks placed into the source code of executable binary files. This r

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning th…

SafetyDGX agent

// Reusable Context Engineering // Context bloat quietly kills long-horizon runs, but you can fix it from the outside without fine-tuning the underlying agent. (bookmark this) Context management is us

SERA: Soft-Verified Efficient Repository Agents

Model ReleasesDGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

Shaft-integrated Force Sensing with Transformer-based Dynamics Compensation for Telesurgery

ApplicationsDGX agent

arXiv:2605.31434v1 Announce Type: new Abstract: Robot-Assisted Minimally Invasive Surgery (RAMIS) enhances surgeon dexterity, with newer platforms leveraging haptic feedback to further improve perform

The Refutability Gap: Challenges in Validating Reasoning by Large Language Models

SafetyDGX agent

arXiv:2601.02380v4 Announce Type: replace-cross Abstract: Recent reports claim that Large Language Models (LLMs) have achieved the ability to derive new science and exhibit human-level general intelli

31 May 2026

OpenAI Robotics is hiring, looking for exceptional full-stack hardware, ops, systems, and ML engineers to help us program and manufacture ro…

TutorialsDGX agent

OpenAI Robotics is hiring, looking for exceptional full-stack hardware, ops, systems, and ML engineers to help us program and manufacture robots that are useful for society. AI should be able to help

29 May 2026

Parallax: Parameterized Local Linear Attention for Language Modeling

Model ReleasesDGX agent

arXiv:2605.29157v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the central paradigm in artificial intelligence, yet the core computational primitive of attention has remain

Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories

Model ReleasesDGX agent

arXiv:2605.29893v1 Announce Type: new Abstract: LLM-based agents have demonstrated strong capabilities in solving complex tasks through multi-step reasoning and tool use. However, existing evaluation

The Open Motion Planning Library 2.0

Model ReleasesDGX agent

arXiv:2605.29301v1 Announce Type: new Abstract: The Open Motion Planning Library (OMPL), first released in 2008, has become a cornerstone of the motion planning community, providing implementations of

28 May 2026

AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models

Model ReleasesDGX agent

arXiv:2602.18481v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has led to a surge of financial benchmarks, evolving from static knowledge evaluation to

Are Large Pre-trained Vision Language Models Effective Construction Safety Inspectors?

Model ReleasesDGX agent

arXiv:2508.11011v2 Announce Type: replace Abstract: Construction safety inspections typically involve a human inspector identifying safety concerns on-site. With the rise of powerful Vision Language M

CFDTwin: An open-source GUI and Python toolkit for POD-NN surrogate modeling of ANSYS Fluent simulations

Model ReleasesDGX agent

arXiv:2605.27725v1 Announce Type: cross Abstract: High-fidelity computational fluid dynamics (CFD) is widely used for thermal-fluid design, but repeated CFD solves remain expensive for design optimiza

Dr-CiK: A Testbed for Foresight-Driven Agents

Model ReleasesDGX agent

arXiv:2605.27904v1 Announce Type: new Abstract: Time series forecasting in real-world settings often depends not only on historical observations, but also on external context that must be actively dis

RASR: Retrieval-Augmented Super Resolution for Practical Reference-based Image Restoration

Model ReleasesDGX agent

arXiv:2508.09449v2 Announce Type: replace Abstract: Reference-based Super Resolution (RefSR) improves upon Single Image Super Resolution (SISR) by leveraging high-quality reference images to enhance t

SAM-Enhanced Segmentation on Road Datasets: Balancing Critical Classes in Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.28136v1 Announce Type: new Abstract: Dense semantic segmentation is essential for autonomous driving, yet many multi-modal datasets lack pixel-level annotations. The Zenseact Open Dataset (

Tackling Multimodal Learning Challenges with Mixture-of-Expert: A Survey

Model ReleasesDGX agent

arXiv:2605.27431v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) presents a naturally compatible and scalable framework for multimodal learning, demonstrating strong adaptability across dive

27 May 2026

Beyond Self-Talk: A Communication-Centric Survey of LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2502.14321v3 Announce Type: replace-cross Abstract: Large language model-based multi-agent systems have recently gained significant attention due to their potential for complex, collaborative, a

Datasets for Lane Detection in Autonomous Driving: A Comprehensive Review

AgentsDGX agent

arXiv:2504.08540v2 Announce Type: replace Abstract: Accurate lane detection is essential for automated driving, enabling safe and reliable vehicle navigation across a variety of road scenarios. Numero

← Previous
1…364365366367368…428
Next →