AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction

DGX agent

arXiv:2604.28115v1 Announce Type: cross Abstract: Existing learning-based occupancy prediction methods rely on large-scale 3D annotations and generalize poorly across environments. We present FreeOcc,

model-releasesarxiv-cs-cv
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Function-based Parametric Co-Design Optimization of Dexterous Hands

DGX agent

arXiv:2604.27557v1 Announce Type: new Abstract: Despite advances in dexterous hand manipulation, robotic hand design is still largely decoupled from task-driven evaluation and control, limiting system

model-releasesarxiv-cs-ro
1 May 2026
Safety

GAVEL: Towards Rule-Based Safety Through Activation Monitoring

DGX agent

arXiv:2601.19768v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly paired with activation-based monitoring to detect and prevent harmful behaviors that may not be appare

safetyarxiv-cs-ai
1 May 2026
Model Releases

Global Optimality for Constrained Exploration via Penalty Regularization

DGX agent

arXiv:2604.28144v1 Announce Type: new Abstract: Efficient exploration is a central problem in reinforcement learning and is often formalized as maximizing the entropy of the state-action occupancy mea

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

GuideDog: A Real-World Egocentric Multimodal Dataset for Blind and Low-Vision Accessibility-Aware Guidance

DGX agent

arXiv:2503.12844v2 Announce Type: replace Abstract: For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individ

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

HQ-UNet: A Hybrid Quantum-Classical U-Net with a Quantum Bottleneck for Remote Sensing Image Segmentation

DGX agent

arXiv:2604.27206v1 Announce Type: new Abstract: Semantic segmentation in remote sensing is commonly addressed using classical deep learning architectures such as U-Net, which require a large number of

model-releasesarxiv-cs-cv
1 May 2026
Safety

Implicit bias produces neural scaling laws in learning curves, from perceptrons to deep networks

DGX agent

arXiv:2505.13230v3 Announce Type: replace Abstract: Scaling laws in deep learning -- empirical power-law relationships linking model performance to resource growth -- have emerged as simple yet striki

safetyarxiv-cs-lg
1 May 2026
Model Releases

Improving Graph Few-shot Learning with Hyperbolic Space and Denoising Diffusion

DGX agent

arXiv:2604.27462v1 Announce Type: cross Abstract: Graph few-shot learning, which focuses on effectively learning from only a small number of labeled nodes to quickly adapt to new tasks, has garnered s

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Iterative Multimodal Retrieval-Augmented Generation for Medical Question Answering

DGX agent

arXiv:2604.27724v1 Announce Type: new Abstract: Medical retrieval-augmented generation (RAG) systems typically operate on text chunks extracted from biomedical literature, discarding the rich visual c

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

JI-ADF: Joint-Individual Learning with Adaptive Decision Fusion for Multimodal Skin Lesion Classification

DGX agent

arXiv:2604.27343v1 Announce Type: new Abstract: Skin lesion classification is essential for early dermatological diagnosis, yet many existing computer-aided systems rely primarily on dermoscopic image

model-releasesarxiv-cs-cv
1 May 2026
Research

Junk DNA Hypothesis: Pruning Small Pre-Trained Weights Irreversibly and Monotonically Impairs 'Difficult' Downstream Tasks in LLMs

DGX agent

arXiv:2310.02277v4 Announce Type: replace-cross Abstract: We present Junk DNA Hypothesis by adopting a novel task-centric angle for the pre-trained weights of large language models (LLMs). It has been

researcharxiv-cs-ai
1 May 2026
Model Releases

Learning Generalizable Multimodal Representations for Software Vulnerability Detection

DGX agent

arXiv:2604.25711v2 Announce Type: replace-cross Abstract: Source code and its accompanying comments are complementary yet naturally aligned modalities-code encodes structural logic while comments capt

model-releasesarxiv-cs-ai
1 May 2026
Research

Learning to Aggregate Zero-Shot LLM Agents for Corporate Disclosure Classification

DGX agent

arXiv:2603.20965v2 Announce Type: replace-cross Abstract: This paper studies whether a lightweight supervised aggregator can combine diverse zero-shot large language model outputs into a stronger down

researcharxiv-cs-ai
1 May 2026
Research

Leveraging Quantum-Based Architectures for Robust Diagnostics

DGX agent

arXiv:2511.12386v2 Announce Type: replace Abstract: Quantum machine learning has emerged as a promising approach for medical image analysis, particularly in settings where compact models and expressiv

researcharxiv-cs-cv
1 May 2026
Research

LLMs as ASP Programmers: Self-Correction Enables Task-Agnostic Nonmonotonic Reasoning

DGX agent

arXiv:2604.27960v1 Announce Type: new Abstract: Recent large language models (LLMs) have achieved impressive reasoning milestones but continue to struggle with high computational costs, logical incons

researcharxiv-cs-ai
1 May 2026
Research

LLMs Capture Emotion Labels, Not Emotion Uncertainty: Distributional Analysis and Calibration of Human--LLM Judgment Gaps

DGX agent

arXiv:2604.27345v1 Announce Type: new Abstract: Human annotators frequently disagree on emotion labels, yet most evaluations of Large Language Model (LLM) emotion annotation collapse these judgments i

researcharxiv-cs-cl
1 May 2026
Model Releases

Math Education Digital Shadows for facilitating learning with LLMs: Math performance, anxiety and confidence in simulated students and AIs

DGX agent

arXiv:2604.27618v1 Announce Type: new Abstract: To enhance LLMs' impact on math education, we need data on their mathematical prowess and biases across prompts. To fill this gap, we introduce MEDS (Ma

model-releasesarxiv-cs-ai
1 May 2026
Safety

Mind the Gap: Structure-Aware Consistency in Preference Learning

DGX agent

arXiv:2604.27733v1 Announce Type: new Abstract: Preference learning has become the foundation of aligning Large Language Models (LLMs) with human intent. Popular methods, such as Direct Preference Opt

safetyarxiv-cs-lg
1 May 2026
Model Releases

MultiBLiMP 1.0: A Massively Multilingual Benchmark of Linguistic Minimal Pairs

DGX agent

arXiv:2504.02768v4 Announce Type: replace Abstract: We introduce MultiBLiMP 1.0, a massively multilingual benchmark of linguistic minimal pairs, covering 101 languages and 2 types of subject-verb agre

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

DGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence Chains

DGX agent

arXiv:2604.27852v1 Announce Type: cross Abstract: Although precise recall is a core objective in Retrieval-Augmented Generation (RAG), a critical oversight persists in the field: improvements in retri

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

DGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

model-releasesarxiv-cs-ai
1 May 2026
Local Ai

OmniDrive-R1: Reinforcement-driven Interleaved Multi-modal Chain-of-Thought for Trustworthy Vision-Language Autonomous Driving

DGX agent

arXiv:2512.14044v3 Announce Type: replace-cross Abstract: The deployment of Vision-Language Models (VLMs) in safety-critical domains like autonomous driving (AD) is critically hindered by reliability

local-aiarxiv-cs-ai
1 May 2026
Applications

Probabilistic Circuits for Irregular Multivariate Time Series Forecasting

DGX agent

arXiv:2604.27814v1 Announce Type: new Abstract: Joint probabilistic modeling is essential for forecasting irregular multivariate time series (IMTS) to accurately quantify uncertainty. Existing approac

applicationsarxiv-cs-lg
1 May 2026
Model Releases

RoadMapper: A Multi-Agent System for Roadmap Generation of Solving Complex Research Problems

DGX agent

arXiv:2604.27616v1 Announce Type: new Abstract: People commonly leverage structured content to accelerate knowledge acquisition and research problem solving. Among these, roadmaps guide researchers th

model-releasesarxiv-cs-cl
1 May 2026
Applications

Robust Learning on Heterogeneous Graphs with Heterophily: A Graph Structure Learning Approach

DGX agent

arXiv:2604.27387v1 Announce Type: new Abstract: Heterogeneous graphs with heterophily have emerged as a powerful abstraction for modeling complex real-world systems, where nodes of different types and

applicationsarxiv-cs-ai
1 May 2026
Agents

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation

DGX agent

arXiv:2604.27620v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) aims to enable an embodied agent to follow natural-language instructions and navigate to a target location in unsee

agentsarxiv-cs-cv
1 May 2026
Safety

Stable Behavior, Limited Variation: Persona Validity in LLM Agents for Urban Sentiment Perception

DGX agent

arXiv:2604.28048v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used as proxies for human perception in urban analysis, yet it remains unclear whether persona prompting p

safetyarxiv-cs-cl
1 May 2026
Hardware

Strait: Perceiving Priority and Interference in ML Inference Serving

DGX agent

arXiv:2604.28175v1 Announce Type: new Abstract: Machine learning (ML) inference serving systems host deep neural network (DNN) models and schedule incoming inference requests across deployed GPUs. How

hardwarearxiv-cs-lg
1 May 2026
Research

Student Classroom Behavior Recognition Based on Improved YOLOv8s

DGX agent

arXiv:2604.27293v1 Announce Type: new Abstract: In classroom teaching, student behavior can reflect their learning state and classroom participation, which is of great significance for teaching qualit

researcharxiv-cs-cv
1 May 2026
Local Ai

TeD-Loc: Text Distillation for Weakly Supervised Object Localization

DGX agent

arXiv:2501.12632v2 Announce Type: replace Abstract: Weakly supervised object localization (WSOL) models are trained using only image-level class labels. They can predict both the object class and spat

local-aiarxiv-cs-cv
1 May 2026
Model Releases

The Inverse-Wisdom Law: Architectural Tribalism and the Consensus Paradox in Agentic Swarms

DGX agent

arXiv:2604.27274v1 Announce Type: new Abstract: As AI transitions toward multi-agent systems (MAS) to solve complex workflows, research paradigms operate on the axiomatic assumption that agent collabo

model-releasesarxiv-cs-ai
1 May 2026
Research

To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing

DGX agent

arXiv:2604.27296v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code editing, yet the prevalent full-code generation paradigm suffers from severe efficiency bo

researcharxiv-cs-cl
1 May 2026
Model Releases

TransVLM: A Vision-Language Framework and Benchmark for Detecting Any Shot Transitions

DGX agent

arXiv:2604.27975v1 Announce Type: cross Abstract: Traditional Shot Boundary Detection (SBD) inherently struggles with complex transitions by formulating the task around isolated cut points, frequently

model-releasesarxiv-cs-ai
1 May 2026
Research

Validating the Clinical Utility of CineECG 3D Reconstructions through Cross-Modal Feature Attribution

DGX agent

arXiv:2604.27017v1 Announce Type: cross Abstract: Deep learning models for 12-lead electrocardiogram (ECG) analysis achieve high diagnostic performance but lack the intuitive interpretability required

researcharxiv-cs-lg
1 May 2026
Research

VTBench: A Multimodal Framework for Time-Series Classification with Chart-Based Representations

DGX agent

arXiv:2604.27259v1 Announce Type: new Abstract: Time-series classification (TSC) has advanced significantly with deep learning, yet most models rely solely on raw numerical inputs, overlooking alterna

researcharxiv-cs-cv
1 May 2026
Model Releases

WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments

DGX agent

arXiv:2604.27776v1 Announce Type: new Abstract: While GUI agents have shown impressive capabilities in common computer-use tasks such as OSWorld, current benchmarks mainly focus on isolated and single

model-releasesarxiv-cs-ai
1 May 2026
Hardware

ZipCCL: Efficient Lossless Data Compression of Communication Collectives for Accelerating LLM Training

DGX agent

arXiv:2604.27844v1 Announce Type: cross Abstract: Communication has emerged as a critical bottleneck in the distributed training of large language models (LLMs). While numerous approaches have been pr

hardwarearxiv-cs-cl
1 May 2026
Model Releases

3D-LENS: A 3D Lifting-based Elevated Novel-view Synthesis method for Single-View Aerial-Ground Re-Identification

DGX agent

arXiv:2604.26520v1 Announce Type: new Abstract: Aerial-Ground Re-Identification (AG-ReID) is constrained by the viewpoint-domain gap, as drastic viewpoint disparities occlude or distort discriminative

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

A Multistage Extraction Pipeline for Long Scanned Financial Documents: An Empirical Study in Industrial KYC Workflows

DGX agent

arXiv:2604.26462v1 Announce Type: new Abstract: Structured information extraction from long, multilingual scanned financial documents is a core requirement in industrial KYC and compliance workflows.

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

A Scoping Review of LLM-as-a-Judge in Healthcare and the MedJUDGE Framework

DGX agent

arXiv:2604.25933v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly generate and process clinical text, scalable evaluation has become critical. LLM-as-a-Judge (LaaJ), which

safetyarxiv-cs-ai
30 Apr 2026
Safety

Accelerating RL Post-Training Rollouts via System-Integrated Speculative Decoding

DGX agent

arXiv:2604.26779v1 Announce Type: cross Abstract: RL post-training of frontier language models is increasingly bottlenecked by autoregressive rollout generation, making rollout acceleration a central

safetyarxiv-cs-cl
30 Apr 2026
Model Releases

Adaptive and Fine-grained Module-wise Expert Pruning for Efficient LoRA-MoE Fine-Tuning

DGX agent

arXiv:2604.26340v1 Announce Type: new Abstract: LoRA-MoE has emerged as an effective paradigm for parameter-efficient fine-tuning, combining the low training cost of LoRA with the increased adaptation

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision

DGX agent

arXiv:2604.26567v1 Announce Type: new Abstract: Despite the rapid progress in data-driven 3D vision, aerial geometric 3D vision remains a formidable challenge due to the severe scarcity of large-scale

model-releasesarxiv-cs-cv
30 Apr 2026
Safety

Beyond Shortcuts: Mitigating Visual Illusions in Frozen VLMs via Qualitative Reasoning

DGX agent

arXiv:2604.26250v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved state-of-the-art performance in general visual tasks, their perceptual robustness remains remarkably b

safetyarxiv-cs-cv
30 Apr 2026
Applications

CheXthought: A global multimodal dataset of clinical chain-of-thought reasoning and visual attention for chest X-ray interpretation

DGX agent

arXiv:2604.26288v1 Announce Type: cross Abstract: Chest X-ray interpretation is one of the most frequently performed diagnostic tasks in medicine and a primary target for AI development, yet current v

applicationsarxiv-cs-ai
30 Apr 2026
Model Releases

ChinaTravel: An Open-Ended Travel Planning Benchmark with Compositional Constraint Validation for Language Agents

DGX agent

arXiv:2412.13682v5 Announce Type: replace Abstract: Travel planning stands out among real-world applications of Language Agents because it couples significant practical demand with a rigorous constrai

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

ClawGym: A Scalable Framework for Building Effective Claw Agents

DGX agent

arXiv:2604.26904v1 Announce Type: cross Abstract: Claw-style environments support multi-step workflows over local files, tools, and persistent workspace states. However, scalable development around th

model-releasesarxiv-cs-ai
30 Apr 2026
← Previous
1…598599600601602…1082
Next →