AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
52,354 results
Local Ai

At FullTilt: Real-Time Open-Set 3D Macromolecule Detection Directly from Tilted 2D Projections

DGX agent

arXiv:2604.10766v1 Announce Type: new Abstract: Open-set 3D macromolecule detection in cryogenic electron tomography eliminates the need for target-specific model retraining. However, strict VRAM cons

local-aiarxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Beyond Monologue: Interactive Talking-Listening Avatar Generation with Conversational Audio Context-Aware Kernels

DGX agent

arXiv:2604.10367v1 Announce Type: new Abstract: Audio-driven human video generation has achieved remarkable success in monologue scenarios, largely driven by advancements in powerful video generation

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0

DGX agent

arXiv:2604.05767v2 Announce Type: replace-cross Abstract: We present BADAS-2.0, the second generation of our collision anticipation system, building on BADAS-1.0, which showed that fine-tuning V-JEPA2

model-releasesarxiv-cs-cl
14 Apr 2026
Research

BiT-MCTS: A Theme-based Bidirectional MCTS Approach to Chinese Fiction Generation

DGX agent

arXiv:2603.14410v3 Announce Type: replace Abstract: Generating long-form linear fiction from open-ended themes remains a major challenge for large language models, which frequently fail to guarantee g

researcharxiv-cs-cl
14 Apr 2026
Model Releases

BITS Pilani at SemEval-2026 Task 9: Structured Supervised Fine-Tuning with DPO Refinement for Polarization Detection

DGX agent

arXiv:2604.11121v1 Announce Type: new Abstract: The POLAR SemEval-2026 Shared Task aims to detect online polarization and focuses on the classification and identification of multilingual, multicultura

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

BlasBench: An Open Benchmark for Irish Speech Recognition

DGX agent

arXiv:2604.10736v1 Announce Type: new Abstract: No open Irish-specific benchmark compares end-user ASR systems under a shared Irish-aware evaluation protocol. To solve this, we release BlasBench, an o

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Both Ends Count! Just How Good are LLM Agents at 'Text-to-Big SQL'?

DGX agent

arXiv:2602.21480v4 Announce Type: replace-cross Abstract: Text-to-SQL and Big Data are both extensively benchmarked fields, yet there is limited research that evaluates them jointly. In the real world

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Bottleneck Tokens for Unified Multimodal Retrieval

DGX agent

arXiv:2604.11095v1 Announce Type: cross Abstract: Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rel

researcharxiv-cs-ai
14 Apr 2026
Research

Catalog-Native LLM: Speaking Item-ID Dialect with Less Entanglement for Recommendation

DGX agent

arXiv:2510.05125v2 Announce Type: replace Abstract: While collaborative filtering delivers predictive accuracy and efficiency, and Large Language Models (LLMs) enable expressive and generalizable reas

researcharxiv-cs-cl
14 Apr 2026
Model Releases

CocoaBench: Evaluating Unified Digital Agents in the Wild

DGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

CoFusion: Multispectral and Hyperspectral Image Fusion via Spectral Coordinate Attention

DGX agent

arXiv:2604.10584v1 Announce Type: new Abstract: Multispectral and Hyperspectral Image Fusion (MHIF) aims to reconstruct high-resolution images by integrating low-resolution hyperspectral images (LRHSI

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

COMPOSITE-Stem

DGX agent

arXiv:2604.09836v1 Announce Type: new Abstract: AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Context-Aware Semantic Segmentation via Stage-Wise Attention

DGX agent

arXiv:2601.11310v2 Announce Type: replace Abstract: Semantic ultra-high-resolution (UHR) image segmentation is essential in remote sensing applications such as aerial mapping and environmental monitor

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training

DGX agent

arXiv:2507.15640v2 Announce Type: replace-cross Abstract: Continual pre-training on small-scale task-specific data is an effective method for improving large language models in new target fields, yet

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

DDO-RM for LLM Preference Optimization: A Minimal Held-Out Benchmark against DPO

DGX agent

arXiv:2604.11119v1 Announce Type: cross Abstract: This paper reorganizes the current manuscript around the DPO versus DDO-RM preference-optimization project and focuses on two parts: the algorithmic v

model-releasesarxiv-cs-lg
14 Apr 2026
Research

DeepSketcher: Internalizing Visual Manipulation for Multimodal Reasoning

DGX agent

arXiv:2509.25866v2 Announce Type: replace Abstract: The 'thinking with images' paradigm represents a pivotal shift in the reasoning of Vision Language Models (VLMs), moving from text-dominant chain-of

researcharxiv-cs-cv
14 Apr 2026
Model Releases

Differentially Private Verification of Distribution Properties

DGX agent

arXiv:2604.10819v1 Announce Type: cross Abstract: A recent line of work initiated by Chiesa and Gur and further developed by Herman and Rothblum investigates the sample and communication complexity of

model-releasesarxiv-cs-lg
14 Apr 2026
Safety

Diffusion-Based Generative Priors for Efficient Beam Alignment in Directional Networks

DGX agent

arXiv:2604.09653v1 Announce Type: cross Abstract: Beam alignment is a key challenge in directional mmWave and THz systems, where narrow beams require accurate yet low-overhead training. Existing learn

safetyarxiv-cs-ai
14 Apr 2026
Research

Diffusion-CAM: Faithful Visual Explanations for dMLLMs

DGX agent

arXiv:2604.11005v1 Announce Type: new Abstract: While diffusion Multimodal Large Language Models (dMLLMs) have recently achieved remarkable strides in multimodal generation, the development of interpr

researcharxiv-cs-ai
14 Apr 2026
Research

Discourse Diversity in Multi-Turn Empathic Dialogue

DGX agent

arXiv:2604.11742v1 Announce Type: cross Abstract: Large language models (LLMs) produce responses rated as highly empathic in single-turn settings (Ayers et al., 2023; Lee et al., 2024), yet they are a

researcharxiv-cs-ai
14 Apr 2026
Safety

Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations

DGX agent

arXiv:2604.11322v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive capabilities in utilizing external tools. In practice, however, LLMs are often exposed to to

safetyarxiv-cs-ai
14 Apr 2026
Applications

EDUMATH: Generating Standards-aligned Educational Math Word Problems

DGX agent

arXiv:2510.06965v2 Announce Type: replace-cross Abstract: Math word problems (MWPs) are critical K-12 educational tools, and customizing them to students' interests and ability levels can enhance lear

applicationsarxiv-cs-ai
14 Apr 2026
Safety

Endogenous Information in Routing Games: Memory-Constrained Equilibria, Recall Braess Paradoxes, and Memory Design

DGX agent

arXiv:2604.11733v1 Announce Type: cross Abstract: We study routing games in which travelers optimize over routes that are remembered or surfaced, rather than over a fixed exogenous action set. The pap

safetyarxiv-cs-ai
14 Apr 2026
Tutorials

Engineering Resource-constrained Software Systems with DNN Components: a Concept-based Pruning Approach

DGX agent

arXiv:2604.09988v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are widely used by engineers to solve difficult problems that require predictive modeling from data. However, these models

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

EviRCOD: Evidence-Guided Probabilistic Decoding for Referring Camouflaged Object Detection

DGX agent

arXiv:2604.10894v1 Announce Type: new Abstract: Referring Camouflaged Object Detection (Ref-COD) focuses on segmenting specific camouflaged targets in a query image using category-aligned references.

model-releasesarxiv-cs-cv
14 Apr 2026
Research

EvoESAP: Non-Uniform Expert Pruning for Sparse MoE

DGX agent

arXiv:2603.06003v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts (SMoE) language models achieve strong capability at low per-token compute, yet deployment remains constrained by memory fo

researcharxiv-cs-lg
14 Apr 2026
Model Releases

Exploring Cross-Modal Flows for Few-Shot Learning

DGX agent

arXiv:2510.14543v4 Announce Type: replace Abstract: Aligning features from different modalities, is one of the most fundamental challenges for cross-modal tasks. Although pre-trained vision-language m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Exploring the best way for UAV visual localization under Low-altitude Multi-view Observation Condition: a Benchmark

DGX agent

arXiv:2503.10692v2 Announce Type: replace Abstract: Absolute Visual Localization (AVL) enables an Unmanned Aerial Vehicle (UAV) to determine its position in GNSS-denied environments by establishing ge

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count

DGX agent

arXiv:2604.09689v1 Announce Type: cross Abstract: Machine learning progress has historically prioritized model-centric innovations, yet achievable performance is frequently capped by the intrinsic com

safetyarxiv-cs-ai
14 Apr 2026
Safety

FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

DGX agent

arXiv:2604.10693v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has improved LLM reasoning, but models often generate explanations that appear coherent while containing unfaithful int

safetyarxiv-cs-ai
14 Apr 2026
Safety

FAITH: Factuality Alignment through Integrating Trustworthiness and Honestness

DGX agent

arXiv:2604.10189v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate factually inaccurate content even if they have corresponding knowledge, which critically undermines their reli

safetyarxiv-cs-cl
14 Apr 2026
Research

From Perception to Planning: Evolving Ego-Centric Task-Oriented Spatiotemporal Reasoning via Curriculum Learning

DGX agent

arXiv:2604.10517v1 Announce Type: new Abstract: Modern vision-language models achieve strong performance in static perception, but remain limited in the complex spatiotemporal reasoning required for e

researcharxiv-cs-ai
14 Apr 2026
Research

GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO

DGX agent

arXiv:2601.06767v2 Announce Type: replace-cross Abstract: We present a Bengali mathematical reasoning model called GanitLLM (named after the Bangla word for mathematics, 'Ganit'), together with a new

researcharxiv-cs-ai
14 Apr 2026
Model Releases

GazeVaLM: A Multi-Observer Eye-Tracking Benchmark for Evaluating Clinical Realism in AI-Generated X-Rays

DGX agent

arXiv:2604.11653v1 Announce Type: new Abstract: We introduce GazeVaLM, a public eye-tracking dataset for studying clinical perception during chest radiograph authenticity assessment. The dataset compr

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

GenProve: Learning to Generate Text with Fine-Grained Provenance

DGX agent

arXiv:2601.04932v2 Announce Type: replace Abstract: Large language models (LLM) often hallucinate, and while adding citations is a common solution, it is frequently insufficient for accountability as

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Governed Reasoning for Institutional AI

DGX agent

arXiv:2604.10658v1 Announce Type: new Abstract: Institutional decisions -- regulatory compliance, clinical triage, prior authorization appeal -- require a different AI architecture than general-purpos

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

HeceTokenizer: A Syllable-Based Tokenization Approach for Turkish Retrieval

DGX agent

arXiv:2604.10665v1 Announce Type: new Abstract: HeceTokenizer is a syllable-based tokenizer for Turkish that exploits the deterministic six-pattern phonological structure of the language to construct

model-releasesarxiv-cs-cl
14 Apr 2026
Research

How LLMs Might Think

DGX agent

arXiv:2604.09674v1 Announce Type: new Abstract: Do large language models (LLMs) think? Daniel Stoljar and Zhihe Vincent Zhang have recently developed an argument from rationality for the claim that LL

researcharxiv-cs-ai
14 Apr 2026
Model Releases

How You Ask Matters! Adaptive RAG Robustness to Query Variations

DGX agent

arXiv:2604.10745v1 Announce Type: new Abstract: Adaptive Retrieval-Augmented Generation (RAG) promises accuracy and efficiency by dynamically triggering retrieval only when needed and is widely used i

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

I Can't Believe TTA Is Not Better: When Test-Time Augmentation Hurts Medical Image Classification

DGX agent

arXiv:2604.09697v1 Announce Type: cross Abstract: Test-time augmentation (TTA)--aggregating predictions over multiple augmented copies of a test input--is widely assumed to improve classification accu

model-releasesarxiv-cs-ai
14 Apr 2026
Research

I Walk the Line: Examining the Role of Gestalt Continuity in Object Binding for Vision Transformers

DGX agent

arXiv:2604.09942v1 Announce Type: cross Abstract: Object binding is a foundational process in visual cognition, during which low-level perceptual features are joined into object representations. Bindi

researcharxiv-cs-ai
14 Apr 2026
Model Releases

IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly

DGX agent

arXiv:2604.10409v1 Announce Type: cross Abstract: We introduce IMPACT, a synchronized five-view RGB-D dataset for deployment-oriented industrial procedural understanding, built around real assembly an

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Improving understanding and trust in AI: How users benefit from interval-based counterfactual explanations

DGX agent

arXiv:2604.09573v1 Announce Type: cross Abstract: Experimental user studies evaluating the effectiveness of different subtypes of post-hoc explanations for black-box models are largely nonexistent. Th

researcharxiv-cs-lg
14 Apr 2026
Local Ai

Inertial Magnetic SLAM Systems Using Low-Cost Sensors

DGX agent

arXiv:2512.10128v2 Announce Type: replace Abstract: Spatially inhomogeneous magnetic fields offer a valuable, non-visual information source for positioning. Among systems leveraging this, magnetic fie

local-aiarxiv-cs-ro
14 Apr 2026
Model Releases

Investigating Bias and Fairness in Appearance-based Gaze Estimation

DGX agent

arXiv:2604.10707v1 Announce Type: new Abstract: While appearance-based gaze estimation has achieved significant improvements in accuracy and domain adaptation, the fairness of these systems across dif

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

KL Divergence Between Gaussians: A Step-by-Step Derivation for the Variational Autoencoder Objective

DGX agent

arXiv:2604.11744v1 Announce Type: new Abstract: Kullback-Leibler (KL) divergence is a fundamental concept in information theory that quantifies the discrepancy between two probability distributions. I

tutorialsarxiv-cs-lg
14 Apr 2026
Research

KSDiff: Keyframe-Augmented Speech-Aware Dual-Path Diffusion for Facial Animation

DGX agent

arXiv:2509.20128v2 Announce Type: replace-cross Abstract: Audio-driven facial animation has made significant progress in multimedia applications, with diffusion models showing strong potential for tal

researcharxiv-cs-ai
14 Apr 2026
Applications

Learnable Motion-Focused Tokenization for Effective and Efficient Video Unsupervised Domain Adaptation

DGX agent

arXiv:2604.09955v1 Announce Type: new Abstract: Video Unsupervised Domain Adaptation (VUDA) poses a significant challenge in action recognition, requiring the adaptation of a model from a labeled sour

applicationsarxiv-cs-cv
14 Apr 2026
← Previous
1…519520521522523…1091
Next →