AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Model Releases

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

DGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

model-releasesarxiv-cs-ai
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models

DGX agent

arXiv:2605.08115v1 Announce Type: cross Abstract: Wepresent Alice v1, a 14-billion parameter open-source video generation model that achieves state-of-the-art quality through consistency distillation

model-releasesarxiv-cs-cv
12 May 2026
Agents

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

DGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

agentsarxiv-cs-ai
12 May 2026
Tutorials

Attention-based graph neural networks: a survey

DGX agent

arXiv:2605.08679v1 Announce Type: cross Abstract: Graph neural networks (GNNs) aim to learn well-trained representations in a lower-dimension space for downstream tasks while preserving the topologica

tutorialsarxiv-cs-ai
12 May 2026
Safety

Balancing Efficiency and Fairness in Traffic Light Control through Deep Reinforcement Learning

DGX agent

arXiv:2605.10170v1 Announce Type: new Abstract: Urban traffic congestion presents a significant challenge for modern cities, which impacts mobility and sustainability. Traditional traffic light contro

safetyarxiv-cs-lg
12 May 2026
Model Releases

BenchHAR: Benchmarking Self-Supervised Learning for Generalizable Sensor-based Activity Recognition

DGX agent

arXiv:2605.08296v1 Announce Type: new Abstract: Human Activity Recognition (HAR) from wearable sensors supports broad healthcare and behavior science applications. However, data heterogeneity and the

model-releasesarxiv-cs-cv
12 May 2026
Tutorials

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators

DGX agent

arXiv:2605.09079v1 Announce Type: new Abstract: Despite surpassing human performance across mathematics, coding, and other knowledge-intensive tasks, large language models (LLMs) continue to struggle

tutorialsarxiv-cs-ai
12 May 2026
Hardware

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

DGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

hardwarearxiv-cs-cv
12 May 2026
Model Releases

ChartDiff: A Large-Scale Benchmark for Comprehending Pairs of Charts

DGX agent

arXiv:2603.28902v2 Announce Type: replace Abstract: Charts are central to analytical reasoning, yet existing benchmarks for chart understanding focus almost exclusively on single-chart interpretation

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Computer Use at the Edge of the Statistical Precipice

DGX agent

arXiv:2605.08261v1 Announce Type: cross Abstract: Evaluating Computer Use Agents (CUAs) on interactive environments is fraught with methodological pitfalls that the field has yet to systematically add

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

ConFit v3: Improving Resume-Job Matching with LLM-based Re-Ranking

DGX agent

arXiv:2605.09760v1 Announce Type: new Abstract: A reliable resume-job matching system helps a company find suitable candidates from a pool of resumes and helps a job seeker find relevant jobs from a l

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Coordinates of Capability: A Unified MTMM-Geometric Framework for LLM Evaluation

DGX agent

arXiv:2605.08522v1 Announce Type: new Abstract: The evaluation of Large Language Models (LLMs) faces a critical challenge in construct validity, where fragmented benchmarks and ad hoc metrics frequent

model-releasesarxiv-cs-cl
12 May 2026
Local Ai

EAR: Enhancing Uni-Modal Representations for Weakly Supervised Audio-Visual Video Parsing

DGX agent

arXiv:2605.08723v1 Announce Type: new Abstract: Weakly supervised Audio-Visual Video Parsing (AVVP) aims to recognize and temporally localize audio, visual, and audio-visual events in videos using onl

local-aiarxiv-cs-cv
12 May 2026
Model Releases

Efficient Evaluation of LLM Performance with Statistical Guarantees

DGX agent

arXiv:2601.20251v3 Announce Type: replace-cross Abstract: Exhaustively evaluating many large language models (LLMs) on a large suite of benchmarks is expensive. We cast benchmarking as finite-populati

model-releasesarxiv-cs-lg
12 May 2026
Local Ai

End-to-End Keyword Spotting on FPGA Using Graph Neural Networks with a Neuromorphic Auditory Sensor

DGX agent

arXiv:2605.09570v1 Announce Type: new Abstract: With the rapid growth of mobile robotics and embedded intelligence, there is an increasing demand for efficient on-device data processing on edge platfo

local-aiarxiv-cs-lg
12 May 2026
Hardware

Energy Consumption of Dataframe Libraries for End-to-End Deep Learning Pipelines:A Comparative Analysis

DGX agent

arXiv:2511.08644v3 Announce Type: replace-cross Abstract: This paper presents a detailed comparative analysis of the performance of three major Python data manipulation libraries - Pandas, Polars, and

hardwarearxiv-cs-ai
12 May 2026
Agents

Engineering Robustness into Personal Agents with the AI Workflow Store

DGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

agentsarxiv-cs-ai
12 May 2026
Model Releases

EpiGraph: A Knowledge Graph and Benchmark for Evidence-Intensive Reasoning in Epilepsy

DGX agent

arXiv:2605.09505v1 Announce Type: new Abstract: Epilepsy diagnosis and treatment require evidence-intensive reasoning across heterogeneous clinical knowledge, including biosignal patterns, genetic mec

model-releasesarxiv-cs-ai
12 May 2026
Agents

Evolutionary Ensemble of Agents

DGX agent

arXiv:2605.09018v1 Announce Type: cross Abstract: We introduce Evolutionary Ensemble (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving sys

agentsarxiv-cs-ai
12 May 2026
Applications

Explainable Machine Learning Framework for Cardiovascular Disease Diagnosis and Prognosis

DGX agent

arXiv:2507.11185v2 Announce Type: replace-cross Abstract: Heart disease continues to pose a critical worldwide health issue, more specifically in areas with insufficient access to healthcare infrastru

applicationsarxiv-cs-ai
12 May 2026
Safety

FairHealth: An Open-Source Python Library for Trustworthy Healthcare AI in Low-Resource Settings

DGX agent

arXiv:2605.08198v1 Announce Type: cross Abstract: We present FairHealth, an open-source Python library that provides a unified, modular framework for trustworthy machine learning in healthcare applica

safetyarxiv-cs-ai
12 May 2026
Safety

Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems

DGX agent

arXiv:2605.10604v1 Announce Type: cross Abstract: Designing fair algorithmic decision systems requires balancing model performance with fairness toward affected individuals: More fairness might requir

safetyarxiv-cs-ai
12 May 2026
Applications

Forecasting Source Stability in Scientific Experiments using Temporal Learning Models: A Case Study from Tritium Monitoring

DGX agent

arXiv:2605.08140v1 Announce Type: cross Abstract: The Karlsruhe Tritium Neutrino Experiment (KATRIN) aims to measure the absolute neutrino mass with unprecedented sensitivity, requiring precise monito

applicationsarxiv-cs-ai
12 May 2026
Model Releases

FormalRewardBench: A Benchmark for Formal Theorem Proving Reward Models

DGX agent

arXiv:2605.10141v1 Announce Type: new Abstract: Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. Whil

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

From Traditional Taggers to LLMs: A Comparative Study of POS Tagging for Medieval Romance Languages

DGX agent

arXiv:2605.09147v1 Announce Type: cross Abstract: Part-of-speech (POS) tagging for Medieval Romance languages remains challenging due to orthographic variation, morphological complexity, and limited a

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

General Agent Evaluation

DGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

GLiNER2-PII: A Multilingual Model for Personally Identifiable Information Extraction

DGX agent

arXiv:2605.09973v1 Announce Type: cross Abstract: Reliable detection of personally identifiable information (PII) is increasingly important across modern data-processing systems, yet the task remains

model-releasesarxiv-cs-ai
12 May 2026
Safety

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

DGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

safetyarxiv-cs-ai
12 May 2026
Model Releases

GraphBench: Next-generation graph learning benchmarking

DGX agent

arXiv:2512.04475v5 Announce Type: replace-cross Abstract: Machine learning on graphs has made substantial progress across domains such as molecular property prediction and chip design. Yet benchmarkin

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

GravityGraphSAGE: Link Prediction in Directed Attributed Graphs

DGX agent

arXiv:2605.09408v1 Announce Type: new Abstract: Link prediction (inferring missing or future connections between nodes in a graph) is a fundamental problem in network science with widespread applicati

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning

DGX agent

arXiv:2605.10546v1 Announce Type: new Abstract: Pixel-based deep reinforcement learning agents are typically trained on heavily downsampled visual observations, a convention inherited from early bench

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Incremental Multilingual Text2Cypher with Adapter Combination

DGX agent

arXiv:2601.16097v2 Announce Type: replace Abstract: Large Language Models enable users to access database using natural language interfaces using tools like Text2SQL, Text2SPARQL, and Text2Cypher, whi

model-releasesarxiv-cs-cl
12 May 2026
Local Ai

Large Language Models over Networks: Collaborative Intelligence under Resource Constraints

DGX agent

arXiv:2605.08626v1 Announce Type: cross Abstract: Large language models (LLMs) are transforming society, powering applications from smartphone assistants to autonomous driving. Yet cloud-based LLM ser

local-aiarxiv-cs-lg
12 May 2026
Hardware

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

DGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

hardwarearxiv-cs-ai
12 May 2026
Model Releases

LLaVA-UHD v4: What Makes Efficient Visual Encoding in MLLMs?

DGX agent

arXiv:2605.08985v1 Announce Type: new Abstract: Visual encoding constitutes a major computational bottleneck in Multimodal Large Language Models (MLLMs), especially for high-resolution image inputs. T

model-releasesarxiv-cs-cv
12 May 2026
Tutorials

LLM-Augmented Chemical Synthesis and Design Decision Programs

DGX agent

arXiv:2505.07027v2 Announce Type: replace Abstract: Retrosynthesis, the process of breaking down a target molecule into simpler precursors through a series of valid reactions, stands at the core of or

tutorialsarxiv-cs-ai
12 May 2026
Hardware

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

DGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

hardwarearxiv-cs-lg
12 May 2026
Model Releases

LLMSYS-HPOBench: Hyperparameter Optimization Benchmark Suite for Real-World LLM Systems

DGX agent

arXiv:2605.08305v1 Announce Type: cross Abstract: Large Language Model (LLM) systems have been the frontier of AI in many application domains, leading to new challenges and opportunities for hyperpara

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Locking Pretrained Weights via Deep Low-Rank Residual Distillation

DGX agent

arXiv:2605.10777v1 Announce Type: new Abstract: The quality of open-weight language models has dramatically improved in recent years. Sharing weights greatly facilitates model adoption by enabling the

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Magis-Bench: Evaluating LLMs on Magistrate-Level Legal Tasks

DGX agent

arXiv:2605.08437v1 Announce Type: cross Abstract: Existing benchmarks for legal AI focus primarily on tasks where LLMs must produce legal arguments or documents, yet the capacity to judge such argumen

model-releasesarxiv-cs-ai
12 May 2026
Agents

MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation

DGX agent

arXiv:2502.13451v5 Announce Type: replace Abstract: Vision-and-language navigation (VLN) is a key task in Embodied AI, requiring agents to navigate diverse and unseen environments while following natu

agentsarxiv-cs-ro
12 May 2026
Model Releases

MemQ: Integrating Q-Learning into Self-Evolving Memory Agents over Provenance DAGs

DGX agent

arXiv:2605.08374v1 Announce Type: new Abstract: Episodic memory allows LLM agents to accumulate and retrieve experience, but current methods treat each memory independently, i.e., evaluating retrieval

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

MulTaBench: Benchmarking Multimodal Tabular Learning with Text and Image

DGX agent

arXiv:2605.10616v1 Announce Type: cross Abstract: Tabular Foundation Models have recently established the state of the art in supervised tabular learning, by leveraging pretraining to learn generaliza

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

Multi-domain Multi-modal Document Classification Benchmark with a Multi-level Taxonomy

DGX agent

arXiv:2605.10550v1 Announce Type: new Abstract: Document classification forms the backbone of modern enterprise content management, yet existing benchmarks remain trapped in oversimplified paradigms -

model-releasesarxiv-cs-cl
12 May 2026
Safety

Navigating LLM Valley: From AdamW to Memory-Efficient and Matrix-Based Optimizers

DGX agent

arXiv:2605.09176v1 Announce Type: cross Abstract: Training large language models requires optimization algorithms that are not only statistically effective, but also computationally and memory efficie

safetyarxiv-cs-ai
12 May 2026
Local Ai

On Distinguishing Capability Elicitation from Capability Creation in Post-Training: A Free-Energy Perspective

DGX agent

arXiv:2605.08368v1 Announce Type: new Abstract: Debates about large language model post-training often treat supervised fine-tuning (SFT) as imitation and reinforcement learning (RL) as discovery. But

local-aiarxiv-cs-ai
12 May 2026
Model Releases

OrderFusion: Encoding Orderbook for End-to-End Probabilistic Intraday Electricity Price Forecasting

DGX agent

arXiv:2502.06830v5 Announce Type: replace-cross Abstract: Probabilistic intraday electricity price forecasting is becoming increasingly important for short-term power-system operation. With increasing

model-releasesarxiv-cs-ai
12 May 2026
Safety

Overcoming Catastrophic Forgetting in Visual Continual Learning with Reinforcement Fine-Tuning

DGX agent

arXiv:2605.09640v1 Announce Type: new Abstract: Recent studies suggest that Reinforcement Fine-Tuning (RFT) is inherently more resilient to catastrophic forgetting than Supervised Fine-Tuning (SFT). H

safetyarxiv-cs-cv
12 May 2026
← Previous
1…441442443444445…462
Next →