AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
59,851 results
Applications

Bayesian Inverse Transition Learning: Learning Dynamics From Near-Optimal Trajectories

DGX agent

arXiv:2411.05174v2 Announce Type: replace Abstract: We consider the problem of estimating the transition dynamics T^* from near-optimal expert trajectories in the context of offline model-based reinfo

applicationsarxiv-cs-lg
29 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Below-Chance Blindness: Prompted Underperformance in Small LLMs Produces Positional Bias Rather than Answer Avoidance

DGX agent

arXiv:2604.25249v1 Announce Type: new Abstract: Detecting sandbagging--the deliberate underperformance on capability evaluations--is an open problem in AI safety. We tested whether symptom validity te

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks

DGX agent

arXiv:2604.24955v1 Announce Type: new Abstract: As benchmarks grow in complexity, many apparent agent failures are not failures of the agent at all - they are failures of the benchmark itself: broken

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking and Adapting On-Device LLMs for Clinical Decision Support

DGX agent

arXiv:2601.03266v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly advanced in clinical decision-making, yet the deployment of proprietary systems is hindered by privacy con

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking and Improving GUI Agents in High-Dynamic Environments

DGX agent

arXiv:2604.25380v1 Announce Type: new Abstract: Recent advancements in Graphical User Interface (GUI) agents have predominantly focused on training paradigms like supervised fine-tuning (SFT) and rein

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Benchmarking Layout-Guided Diffusion Models through Unified Semantic-Spatial Evaluation in Closed and Open Settings

DGX agent

arXiv:2604.25358v1 Announce Type: new Abstract: Evaluating layout-guided text-to-image generative models requires assessing both semantic alignment with textual prompts and spatial fidelity to prescri

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Benchmarking Logistic Regression, SVM, and LightGBM Against BiLSTM with Attention for Sentiment Analysis on Indonesian Product Reviews

DGX agent

arXiv:2604.25452v1 Announce Type: new Abstract: Sentiment analysis of product reviews on e-commerce platforms plays a critical role in automatically understanding customer satisfaction and providing a

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Benchmarking OCR Pipelines with Adaptive Enhancement for Multi-Domain Retail Bill Digitization

DGX agent

arXiv:2604.25176v1 Announce Type: new Abstract: The digitization of multi-domain retail billing documents remains a challenging task due to variability in scan quality, layout heterogeneity, and domai

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Benchmarking PyCaret AutoML Against IndoBERT Fine-Tuning for Sentiment Analysis on Indonesian IKN Twitter Data

DGX agent

arXiv:2604.25392v1 Announce Type: new Abstract: This paper benchmarks a classical machine learning approach based on PyCaret AutoML against a deep learning approach based on IndoBERT fine-tuning for b

researcharxiv-cs-cl
29 Apr 2026
Agents

BEVal: A Cross-dataset Evaluation Study of BEV Segmentation Models for Autonomous Driving

DGX agent

arXiv:2408.16322v4 Announce Type: replace Abstract: Current research in semantic bird's-eye view segmentation for autonomous driving focuses solely on optimizing neural network models using a single d

agentsarxiv-cs-cv
29 Apr 2026
Safety

Beyond Accuracy: Benchmarking Cross-Task Consistency in Unified Multimodal Models

DGX agent

arXiv:2604.25072v1 Announce Type: new Abstract: Unified Multimodal Models (uMMs) aim to support both visual understanding and visual generation within a shared representation. However, existing evalua

safetyarxiv-cs-cv
29 Apr 2026
Research

Beyond Fidelity: Semantic Similarity Assessment in Low-Level Image Processing

DGX agent

arXiv:2604.25408v1 Announce Type: new Abstract: Low-level image processing has long been evaluated mainly from the perspective of visual fidelity. However, with the rise of deep learning and generativ

researcharxiv-cs-cv
29 Apr 2026
Model Releases

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal

DGX agent

arXiv:2509.09708v3 Announce Type: replace Abstract: Refusal on harmful prompts is a key safety behaviour in instruction-tuned large language models (LLMs), yet the internal causes of this behaviour re

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Biased Dreams: Limitations to Epistemic Uncertainty Quantification in Latent Space Models

DGX agent

arXiv:2604.25416v1 Announce Type: new Abstract: Model-Based Reinforcement Learning distinguishes between physical dynamics models operating on proprioceptive inputs and latent dynamics models operatin

researcharxiv-cs-lg
29 Apr 2026
Model Releases

BifDet: A 3D Bifurcation Detection Dataset for Airway-Tree Modeling

DGX agent

arXiv:2604.24999v1 Announce Type: new Abstract: Thoracic Computed Tomography (CT) scans offer detailed insights into the intricate branching network of the airway tree, which is essential for understa

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

BLASST: Dynamic BLocked Attention Sparsity via Softmax Thresholding

DGX agent

arXiv:2512.12087v3 Announce Type: replace Abstract: The growing demand for long-context inference capabilities in Large Language Models (LLMs) has intensified the computational and memory bottlenecks

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Bridging the Indoor-Outdoor Gap: Cross-Technology Ranging for Seamless Robot Navigation

DGX agent

arXiv:2604.25541v1 Announce Type: cross Abstract: Mobile robots that move between outdoor and indoor environments still struggle with consistent positioning. Satellite-based and terrestrial ranging ea

researcharxiv-cs-ro
29 Apr 2026
Local Ai

Bug-Report-Driven Fault Localization: Industrial Benchmarking and Lesson Learned at ABB Robotics

DGX agent

arXiv:2604.25700v1 Announce Type: cross Abstract: Software quality assurance remains a major challenge in industrial environments, where large-scale and long-lived systems inevitably accumulate defect

local-aiarxiv-cs-lg
29 Apr 2026
Research

Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation

DGX agent

arXiv:2604.25580v1 Announce Type: new Abstract: The closure of Perspective API at the end of 2026 discards what has functioned as the de facto standard for automated toxicity measurement in NLP, CSS,

researcharxiv-cs-cl
29 Apr 2026
Tutorials

C3G: Learning Compact 3D Representations with 2K Gaussians

DGX agent

arXiv:2512.04021v2 Announce Type: replace Abstract: Reconstructing and understanding 3D scenes from unposed sparse views in a feed-forward manner remains as a challenging task in 3D computer vision. R

tutorialsarxiv-cs-cv
29 Apr 2026
Research

Calibrated Fusion for Heterogeneous Graph-Vector Retrieval in Multi-Hop QA

DGX agent

arXiv:2603.28886v2 Announce Type: replace-cross Abstract: Graph-augmented retrieval combines dense similarity with graph-based relevance signals such as Personalized PageRank (PPR), but these scores h

researcharxiv-cs-lg
29 Apr 2026
Model Releases

CAN-QA: A Question-Answering Benchmark for Reasoning over In-Vehicle CAN Traffic

DGX agent

arXiv:2604.24935v1 Announce Type: cross Abstract: The Controller Area Network (CAN) is a safety-critical in-vehicle communication protocol that lacks built-in security mechanisms, making intrusion det

model-releasesarxiv-cs-lg
29 Apr 2026
Research

Can We Change the Stroke Size for Easier Diffusion?

DGX agent

arXiv:2603.26783v2 Announce Type: replace Abstract: Diffusion models can be challenged in the low signal-to-noise regime, where they have to make pixel-level predictions despite the presence of high n

researcharxiv-cs-cv
29 Apr 2026
Safety

Carbon-Taxed Transformers: A Green Compression Pipeline for Overgrown Language Models

DGX agent

arXiv:2604.25903v1 Announce Type: cross Abstract: The accelerating adoption of Large Language Models (LLMs) in software engineering (SE) has brought with it a silent crisis: unsustainable computationa

safetyarxiv-cs-lg
29 Apr 2026
Research

Categorical Optimization with Bayesian Anchored Latent Trust Regions for Structural Design under High-Dimensional Uncertainty

DGX agent

arXiv:2604.25241v1 Announce Type: new Abstract: Categorical structural optimization under aleatoric uncertainty is challenging because each design variable must be selected from a finite catalog of ad

researcharxiv-cs-lg
29 Apr 2026
Model Releases

CGU-ILALab at FoodBench-QA 2026: Comparing Traditional and LLM-based Approaches for Recipe Nutrient Estimation

DGX agent

arXiv:2604.25774v1 Announce Type: new Abstract: Accurate nutrient estimation from unstructured recipe text is an important yet challenging problem in dietary monitoring, due to ambiguous ingredient te

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Cheaper, Better, Faster, Stronger: Robust Text-to-SQL without Chain-of-Thought or Fine-Tuning

DGX agent

arXiv:2505.14174v2 Announce Type: replace Abstract: LLMs are effective at code generation tasks like text-to-SQL, but is it worth the cost? Many state-of-the-art approaches use non-task-specific LLM t

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

CHUCKLE -- When Humans Teach AI To Learn Emotions The Easy Way

DGX agent

arXiv:2510.09382v2 Announce Type: replace Abstract: Curriculum learning (CL) structures training from simple to complex samples, facilitating progressive learning. However, existing CL approaches for

safetyarxiv-cs-lg
29 Apr 2026
Model Releases

Citation Failure: Definition, Analysis and Efficient Mitigation

DGX agent

arXiv:2510.20303v3 Announce Type: replace Abstract: Citations from LLM-based RAG systems are supposed to simplify response verification. However, this goal is undermined in cases of citation failure,

model-releasesarxiv-cs-cl
29 Apr 2026
Research

CiteRadar: A Citation Intelligence Platform for Researcher Profiling and Geographic Visualization

DGX agent

arXiv:2604.25057v1 Announce Type: new Abstract: Understanding the geographic reach and community structure of one's scholarly citations is increasingly valuable for career development, grant applicati

researcharxiv-cs-lg
29 Apr 2026
Research

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding

DGX agent

arXiv:2602.01785v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved remarkable success in source code understanding, yet as software systems grow in scale, computational eff

researcharxiv-cs-cl
29 Apr 2026
Model Releases

Combating Visual Neglect and Semantic Drift in Large Multimodal Models for Enhanced Cross-Modal Retrieval

DGX agent

arXiv:2604.25273v1 Announce Type: new Abstract: Despite significant progress in Unified Multimodal Retrieval (UMR) powered by Large Multimodal Models (LMMs), existing embedding methods primarily focus

model-releasesarxiv-cs-cv
29 Apr 2026
Research

Comparative Study of Bending Analysis using Physics-Informed Neural Networks and Numerical Dynamic Deflection in Perforated nanobeam

DGX agent

arXiv:2604.24768v1 Announce Type: new Abstract: In this chapter, we investigate the bending behavior of a perforated nanobeam subjected to sinusoidal loading using an efficient and computationally rob

researcharxiv-cs-lg
29 Apr 2026
Model Releases

Comparing Data Assimilation and Likelihood-Based Inference on Latent State Estimation in Agent-Based Models

DGX agent

arXiv:2509.17625v2 Announce Type: replace Abstract: In this paper, we present the first systematic comparison of Data Assimilation (DA) and Likelihood-Based Inference (LBI) in the context of an Agent-

model-releasesarxiv-cs-lg
29 Apr 2026
Local Ai

COMPASS: COmpact Multi-channel Prior-map And Scene Signature for Floor-Plan-Based Visual Localization

DGX agent

arXiv:2604.25388v1 Announce Type: new Abstract: Architectural floor plans are widely available priors which contain not only geometry but also the semantic information of the environment, yet existing

local-aiarxiv-cs-cv
29 Apr 2026
Safety

Compute Aligned Training: Optimizing for Test Time Inference

DGX agent

arXiv:2604.24957v1 Announce Type: new Abstract: Scaling test-time compute has emerged as a powerful mechanism for enhancing Large Language Model (LLM) performance. However, standard post-training para

safetyarxiv-cs-lg
29 Apr 2026
Research

Conditional Flow Matching for Probabilistic Downscaling of Maximum 3-day Snowfall in Alaska

DGX agent

arXiv:2604.25172v1 Announce Type: cross Abstract: Precipitation in complex terrain is governed by orographic processes operating at scales of a few kilometers, yet climate models typically run at reso

researcharxiv-cs-lg
29 Apr 2026
Safety

Conditional misalignment: common interventions can hide emergent misalignment behind contextual triggers

DGX agent

arXiv:2604.25891v1 Announce Type: new Abstract: Finetuning a language model can lead to emergent misalignment (EM) [Betley et al., 2025b]. Models trained on a narrow distribution of misaligned behavio

safetyarxiv-cs-lg
29 Apr 2026
Model Releases

Contrast-Enhanced Gating in GRUs for Robust Low-Data Sequence Learning

DGX agent

arXiv:2402.09034v3 Announce Type: replace Abstract: Activation functions govern how recurrent networks regulate and transmit information across temporal dependencies. Despite advances in sequence mode

model-releasesarxiv-cs-lg
29 Apr 2026
Tutorials

Contrastive Image-Metadata Pre-Training for Materials Transmission Electron Microscopy

DGX agent

arXiv:2604.24909v1 Announce Type: new Abstract: The vast majority of transmission electron microscopy (TEM) data never gets published and ends up on a backup drive until deleted to free up space. Thes

tutorialsarxiv-cs-lg
29 Apr 2026
Agents

Control Your Queries: Heterogeneous Query Interaction for Camera-Radar Fusion

DGX agent

arXiv:2604.25574v1 Announce Type: new Abstract: In autonomous driving, camera-radar fusion offers complementary sensing and low deployment cost. Existing methods perform fusion through input mixing, f

agentsarxiv-cs-cv
29 Apr 2026
Agents

Cooperate to Compete: Strategic Coordination in Multi-Agent Conquest

DGX agent

arXiv:2604.25088v1 Announce Type: cross Abstract: Language Model (LM)-based agents remain largely untested in mixed-motive settings where agents must leverage short-term cooperation for long-term comp

agentsarxiv-cs-cl
29 Apr 2026
Safety

CORAL: Adaptive Retrieval Loop for Culturally-Aligned Multilingual RAG

DGX agent

arXiv:2604.25676v1 Announce Type: new Abstract: Multilingual retrieval-augmented generation (mRAG) is often implemented within a fixed retrieval space, typically via query or document translation or m

safetyarxiv-cs-cl
29 Apr 2026
Model Releases

CoRE: Concept-Reasoning Expansion for Continual Brain Lesion Segmentation

DGX agent

arXiv:2604.25376v1 Announce Type: new Abstract: Accurate brain lesion segmentation in MRI is vital for effective clinical diagnosis and treatment planning. Due to high annotation costs and strict data

model-releasesarxiv-cs-cv
29 Apr 2026
Research

CoreFlow: Low-Rank Matrix Generative Models

DGX agent

arXiv:2604.24959v1 Announce Type: new Abstract: Learning matrix-valued distributions from high-dimensional and possibly incomplete training data is challenging: ambient-space generative modeling is co

researcharxiv-cs-lg
29 Apr 2026
Research

Cornserve: A Distributed Serving System for Any-to-Any Multimodal Models

DGX agent

arXiv:2603.12118v2 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of multimodal data (e.g., text, image, video, audio) as input

researcharxiv-cs-lg
29 Apr 2026
Model Releases

CRAFT: Grounded Multi-Agent Coordination Under Partial Information

DGX agent

arXiv:2603.25268v2 Announce Type: replace Abstract: We introduce CRAFT, a multi-agent benchmark for evaluating pragmatic communication in large language models under strict partial information. In thi

model-releasesarxiv-cs-cl
29 Apr 2026
Research

CRC-SAM: SAM-Based Multi-Modal Segmentation and Quantification of Colorectal Cancer in CT, Colonoscopy, and Histology Images

DGX agent

arXiv:2604.24793v1 Announce Type: cross Abstract: We present CRC-SAM, a unified framework for colorectal cancer segmentation across colonoscopy, CT, and histopathology images. Unlike prior single-moda

researcharxiv-cs-cv
29 Apr 2026
← Previous
1…10191020102110221023…1247
Next →