AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,629 results
12 May 2026

Attention-based graph neural networks: a survey

TutorialsDGX agent

arXiv:2605.08679v1 Announce Type: cross Abstract: Graph neural networks (GNNs) aim to learn well-trained representations in a lower-dimension space for downstream tasks while preserving the topologica

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

SafetyDGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition

Model ReleasesDGX agent

arXiv:2605.08296v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports broad healthcare and behavior science applications. However, data heterogeneity and the

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond source code: The files AI coding agents trust — and attackers exploit

Model ReleasesDGX agent

As AI coding agents become deeply embedded in developer workflows, defenders must evolve their definition of malicious files and rethink how to protect against them. Autonomous AI agents operate acros

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators

TutorialsDGX agent

arXiv:2605.09079v1 Announce Type: new Abstract: Despite surpassing human performance across mathematics, coding, and other knowledge-intensive tasks, large language models (LLMs) continue to struggle

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

HardwareDGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

ChartDiff: A Large-Scale Benchmark for Comprehending Pairs of Charts

Model ReleasesDGX agent

arXiv:2603.28902v2 Announce Type: replace Abstract: Charts are central to analytical reasoning, yet existing benchmarks for chart understanding focus almost exclusively on single-chart interpretation

Computer Use at the Edge of the Statistical Precipice

Model ReleasesDGX agent

arXiv:2605.08261v1 Announce Type: cross Abstract: Evaluating Computer Use Agents (CUAs) on interactive environments is fraught with methodological pitfalls that the field has yet to systematically add

ConFit v3: Improving Resume-Job Matching with LLM-based Re-Ranking

Model ReleasesDGX agent

arXiv:2605.09760v1 Announce Type: new Abstract: A reliable resume-job matching system helps a company find suitable candidates from a pool of resumes and helps a job seeker find relevant jobs from a l

Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation

Model ReleasesDGX agent

arXiv:2605.08522v1 Announce Type: new Abstract: The evaluation of Large Language Models (LLMs) faces a critical challenge in construct validity, where fragmented benchmarks and ad hoc metrics frequent

Did Jensen Huang catch conflict of interest disease from Sam?

HardwareDGX agent

Did Jensen Huang catch conflict of interest disease from Sam? HUANG FOUNDATION SIGNS GPU COMPUTE DEAL WITH COREWEAVE $NVDA proxy says the charitable foundation tied to Jensen and Lori Huang entered an

EAR: Enhancing Uni-Modal Representations for Weakly Supervised Audio-Visual Video Parsing

Local AiDGX agent

arXiv:2605.08723v1 Announce Type: new Abstract: Weakly supervised Audio-Visual Video Parsing (AVVP) aims to recognize and temporally localize audio, visual, and audio-visual events in videos using onl

Efficient Evaluation of LLM Performance with Statistical Guarantees

Model ReleasesDGX agent

arXiv:2601.20251v3 Announce Type: replace-cross Abstract: Exhaustively evaluating many large language models (LLMs) on a large suite of benchmarks is expensive. We cast benchmarking as finite-populati

End-to-End Keyword Spotting on FPGA Using Graph Neural Networks with a Neuromorphic Auditory Sensor

Local AiDGX agent

arXiv:2605.09570v1 Announce Type: new Abstract: With the rapid growth of mobile robotics and embedded intelligence, there is an increasing demand for efficient on-device data processing on edge platfo

Energy Consumption of Dataframe Libraries for End-to-End Deep Learning Pipelines:A Comparative Analysis

HardwareDGX agent

arXiv:2511.08644v3 Announce Type: replace-cross Abstract: This paper presents a detailed comparative analysis of the performance of three major Python data manipulation libraries - Pandas, Polars, and

Engineering Robustness into Personal Agents with the AI Workflow Store

AgentsDGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

EpiGraph: A Knowledge Graph and Benchmark for Evidence-Intensive Reasoning in Epilepsy

Model ReleasesDGX agent

arXiv:2605.09505v1 Announce Type: new Abstract: Epilepsy diagnosis and treatment require evidence-intensive reasoning across heterogeneous clinical knowledge, including biosignal patterns, genetic mec

Evolutionary Ensemble of Agents

AgentsDGX agent

arXiv:2605.09018v1 Announce Type: cross Abstract: We introduce Evolutionary Ensemble (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving sys

Explainable Machine Learning Framework for Cardiovascular Disease Diagnosis and Prognosis

ApplicationsDGX agent

arXiv:2507.11185v2 Announce Type: replace-cross Abstract: Heart disease continues to pose a critical worldwide health issue, more specifically in areas with insufficient access to healthcare infrastru

FairHealth: An Open-Source Python Library for Trustworthy Healthcare AI in Low-Resource Settings

SafetyDGX agent

arXiv:2605.08198v1 Announce Type: cross Abstract: We present FairHealth, an open-source Python library that provides a unified, modular framework for trustworthy machine learning in healthcare applica

Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems

SafetyDGX agent

arXiv:2605.10604v1 Announce Type: cross Abstract: Designing fair algorithmic decision systems requires balancing model performance with fairness toward affected individuals: More fairness might requir

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

ApplicationsDGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

FormalRewardBench: A Benchmark for Formal Theorem Proving Reward Models

Model ReleasesDGX agent

arXiv:2605.10141v1 Announce Type: new Abstract: Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. Whil

From Traditional Taggers to LLMs: A Comparative Study of POS Tagging for Medieval Romance Languages

Model ReleasesDGX agent

arXiv:2605.09147v1 Announce Type: cross Abstract: Part-of-speech (POS) tagging for Medieval Romance languages remains challenging due to orthographic variation, morphological complexity, and limited a

General Agent Evaluation

Model ReleasesDGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction

Model ReleasesDGX agent

arXiv:2605.09973v1 Announce Type: cross Abstract: Reliable detection of personally identifiable information (PII) is increasingly important across modern data-processing systems, yet the task remains

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

SafetyDGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

GraphBench: Next-generation graph learning benchmarking

Model ReleasesDGX agent

arXiv:2512.04475v5 Announce Type: replace-cross Abstract: Machine learning on graphs has made substantial progress across domains such as molecular property prediction and chip design. Yet benchmarkin

GravityGraphSAGE: Link Prediction in Directed Attributed Graphs

Model ReleasesDGX agent

arXiv:2605.09408v1 Announce Type: new Abstract: Link prediction (inferring missing or future connections between nodes in a graph) is a fundamental problem in network science with widespread applicati

Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning

Model ReleasesDGX agent

arXiv:2605.10546v1 Announce Type: new Abstract: Pixel-based deep reinforcement learning agents are typically trained on heavily downsampled visual observations, a convention inherited from early bench

Incremental Multilingual Text2Cypher with Adapter Combination

Model ReleasesDGX agent

arXiv:2601.16097v2 Announce Type: replace Abstract: Large Language Models enable users to access database using natural language interfaces using tools like Text2SQL, Text2SPARQL, and Text2Cypher, whi

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

Local AiDGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

HardwareDGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

LLaVA-UHD v4: What Makes Efficient Visual Encoding in MLLMs?

Model ReleasesDGX agent

arXiv:2605.08985v1 Announce Type: new Abstract: Visual encoding constitutes a major computational bottleneck in Multimodal Large Language Models (MLLMs), especially for high-resolution image inputs. T

LLM-Augmented Chemical Synthesis and Design Decision Programs

TutorialsDGX agent

arXiv:2505.07027v2 Announce Type: replace Abstract: Retrosynthesis, the process of breaking down a target molecule into simpler precursors through a series of valid reactions, stands at the core of or

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

HardwareDGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

LLMSYS-HPOBench: Hyperparameter Optimization Benchmark Suite for Real-World LLM Systems

Model ReleasesDGX agent

arXiv:2605.08305v1 Announce Type: cross Abstract: Large Language Model (LLM) systems have been the frontier of AI in many application domains, leading to new challenges and opportunities for hyperpara

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

Model ReleasesDGX agent

arXiv:2605.10777v1 Announce Type: new Abstract: The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling the

Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks

Model ReleasesDGX agent

arXiv:2605.08437v1 Announce Type: cross Abstract: Existing benchmarks for legal AI focus primarily on tasks where LLMs must produce legal arguments or documents, yet the capacity to judge such argumen

MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation

AgentsDGX agent

arXiv:2502.13451v5 Announce Type: replace Abstract: Vision-and-language navigation (VLN) is a key task in Embodied AI, requiring agents to navigate diverse and unseen environments while following natu

Meet physics-intern🧑‍🎓, our agentic framework for theoretical physics. It takes Gemini 3.1 Pro from 17.7% to 31.4% on CritPt, a new SOTA o…

Model ReleasesDGX agent

Meet physics-intern🧑‍🎓, our agentic framework for theoretical physics. It takes Gemini 3.1 Pro from 17.7% to 31.4% on CritPt, a new SOTA on one of the hardest benchmarks for LLMs. Theoretical physics

MemQ: Integrating Q-Learning into Self-Evolving Memory Agents over Provenance DAGs

Model ReleasesDGX agent

arXiv:2605.08374v1 Announce Type: new Abstract: Episodic memory allows LLM agents to accumulate and retrieve experience, but current methods treat each memory independently, i.e., evaluating retrieval

MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image

Model ReleasesDGX agent

arXiv:2605.10616v1 Announce Type: cross Abstract: Tabular Foundation Models have recently established the state of the art in supervised tabular learning, by leveraging pretraining to learn generaliza

Multi-domain Multi-modal Document Classification Benchmark with a Multi-level Taxonomy

Model ReleasesDGX agent

arXiv:2605.10550v1 Announce Type: new Abstract: Document classification forms the backbone of modern enterprise content management, yet existing benchmarks remain trapped in oversimplified paradigms -

Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers

SafetyDGX agent

arXiv:2605.09176v1 Announce Type: cross Abstract: Training large language models requires optimization algorithms that are not only statistically effective, but also computationally and memory efficie

On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective

Local AiDGX agent

arXiv:2605.08368v1 Announce Type: new Abstract: Debates about large language model post-training often treat supervised fine-tuning (SFT) as imitation and reinforcement learning (RL) as discovery. But

OrderFusion: Encoding Orderbook for End-to-End Probabilistic Intraday Electricity Price Forecasting

Model ReleasesDGX agent

arXiv:2502.06830v5 Announce Type: replace-cross Abstract: Probabilistic intraday electricity price forecasting is becoming increasingly important for short-term power-system operation. With increasing

Overcoming Catastrophic Forgetting in Visual Continual Learning with Reinforcement Fine-Tuning

SafetyDGX agent

arXiv:2605.09640v1 Announce Type: new Abstract: Recent studies suggest that Reinforcement Fine-Tuning (RFT) is inherently more resilient to catastrophic forgetting than Supervised Fine-Tuning (SFT). H

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT…

Model ReleasesDGX agent

parameter golf was a blast. 2,000+ submissions. 1,000+ verified github accounts. ideas ranging from quantization and depth recurrence to TTT LoRA, SSMs, H-nets, JEPA, and more. autoresearch made itera

PlantMarkerBench: A Multi-Species Benchmark for Evidence-Grounded Plant Marker Reasoning

Model ReleasesDGX agent

arXiv:2605.10032v1 Announce Type: new Abstract: Cell-type-specific marker genes are fundamental to plant biology, yet existing resources primarily rely on curated databases or high-throughput studies

Playing Games with My Heart: An Evaluation of AI Companion Apps

ApplicationsDGX agent

arXiv:2605.08093v1 Announce Type: cross Abstract: The use of chatbots for various forms of companionship is growing rapidly, raising a myriad of questions about simulated relationships, emotional depe

PolarVSR: A Unified Framework and Benchmark for Continuous Space-Time Polarization Video Reconstruction

Model ReleasesDGX agent

arXiv:2605.10275v1 Announce Type: new Abstract: Polarimetric imaging captures surface polarization characteristics, such as the Degree of Linear Polarization (DoLP) and the Angle of Polarization (AoP)

Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari

Model ReleasesDGX agent

arXiv:2605.08578v1 Announce Type: cross Abstract: Developing generalist systems that retain human-like data efficiency is a central challenge. While world models (WMs) offer a promising path, existing

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

Model ReleasesDGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

AgentsDGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

Sam Altman says Elon Musk’s mind games were damaging OpenAI

IndustryDGX agent

OpenAI CEO Sam Altman says Elon Musk did 'huge damage' to the culture of the AI startup. During testimony as part of Musk's lawsuit against OpenAI, Altman said Musk required OpenAI president Greg Broc

SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation

AgentsDGX agent

arXiv:2602.04712v2 Announce Type: replace-cross Abstract: We present a visual-context image-retrieval-augmented generation (ImageRAG)- assisted AI agent for automatic target recognition (ATR) of synth

Scaling Mobile Agent Systems: From Capability Density to Collective Intelligence

Local AiDGX agent

arXiv:2605.08124v1 Announce Type: cross Abstract: Mobile agent systems are emerging as a key paradigm for enabling intelligent applications on edge devices and in AIoT ecosystems. However, their scala

ScholarPeer: A Context-Aware Multi-Agent Framework for Automated Peer Review

AgentsDGX agent

arXiv:2601.22638v2 Announce Type: replace-cross Abstract: The exponential growth of machine learning submissions has strained the traditional peer review process, resulting in slow feedback loops for

Sens-VisualNews: A Benchmark Dataset for Sensational Image Detection

Model ReleasesDGX agent

arXiv:2605.10394v1 Announce Type: new Abstract: The detection of sensational content in media items can be a critical filtering mechanism for identifying check-worthy content and flagging potential di

← Previous
1…406407408409410…428
Next →