AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Model Releases

HBGSA: Hydrogen Bond Graph with Self-Attention for Drug-Target Binding Affinity Prediction

DGX agent

arXiv:2604.23115v1 Announce Type: new Abstract: Accurate prediction of drug-target binding affinity accelerates drug discovery by prioritizing compounds for experimental validation. Current methods fa

model-releasesarxiv-cs-lg
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs

DGX agent

arXiv:2512.16378v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand beyond text, integrating speech as a native modality has given rise to SpeechLLMs, which directly proce

researcharxiv-cs-ai
28 Apr 2026
Local Ai

Hidden States Know Where Reasoning Diverges: Credit Assignment via Span-Level Wasserstein Distance

DGX agent

arXiv:2604.23318v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) performs coarse-grained credit assignment in reinforcement learning with verifiable rewards (RLVR) by assignin

local-aiarxiv-cs-cl
28 Apr 2026
Model Releases

INHerit-SG: Incremental Hierarchical Semantic Scene Graphs with RAG-Style Retrieval

DGX agent

arXiv:2602.12971v2 Announce Type: replace Abstract: Driven by recent advancements in foundation models, semantic scene graphs have emerged as a promising paradigm for high-level 3D environmental abstr

model-releasesarxiv-cs-ro
28 Apr 2026
Safety

Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens

DGX agent

arXiv:2508.01191v5 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has been shown to be effective in eliciting structured reasoning (i.e., CoT reasoning) from large language models (

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Learning Gradient-based Mixup with Extrapolation toward Flatter Minima for Domain Generalization

DGX agent

arXiv:2209.14742v2 Announce Type: replace Abstract: To address distribution shifts between training and test data, domain generalization (DG) leverages multiple source domains to learn a model that ge

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

DGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

LLMs Reading the Rhythms of Daily Life: Aligned Understanding for Behavior Prediction and Generation

DGX agent

arXiv:2604.23578v1 Announce Type: cross Abstract: Human daily behavior unfolds as complex sequences shaped by intentions, preferences, and context. Effectively modeling these behaviors is crucial for

safetyarxiv-cs-ai
28 Apr 2026
Tutorials

Mitigating Error Amplification in Fast Adversarial Training

DGX agent

arXiv:2604.24332v1 Announce Type: new Abstract: Fast Adversarial Training (FAT) has proven effective in enhancing model robustness by encouraging networks to learn perturbation-invariant representatio

tutorialsarxiv-cs-lg
28 Apr 2026
Model Releases

Multi-View Synergistic Learning with Vision-Language Adaption for Low-Resource Biomedical Image Classification

DGX agent

arXiv:2604.23977v1 Announce Type: new Abstract: Accurate biomedical image classification under low-resource conditions remains challenging due to limited annotations, subtle inter-class visual differe

model-releasesarxiv-cs-cv
28 Apr 2026
Research

MUSIC: Learning Muscle-Driven Dexterous Hand Control

DGX agent

arXiv:2604.23886v1 Announce Type: cross Abstract: We present a data-driven approach for physics-based, muscle-driven dexterous control that enables musculoskeletal hands to perform precise piano playi

researcharxiv-cs-ai
28 Apr 2026
Model Releases

OmniSch: A Multimodal PCB Schematic Benchmark For Structured Diagram Visual Reasoning

DGX agent

arXiv:2604.00270v2 Announce Type: replace Abstract: Recent large multimodal models (LMMs) have made rapid progress in visual grounding, document understanding, and diagram reasoning tasks. However, th

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

DGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Parameter-Efficient Multi-Task Learning via Progressive Task-Specific Adaptation

DGX agent

arXiv:2509.19602v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning methods have emerged as a promising solution for adapting pre-trained models to various downstream tasks. While thes

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

DGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

safetyarxiv-cs-cl
28 Apr 2026
Safety

Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs

DGX agent

arXiv:2604.04177v2 Announce Type: replace Abstract: As large language models (LLMs) are increasing integrated into fact-checking pipelines, formal logic is often proposed as a rigorous means by which

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

DGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

DGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

researcharxiv-cs-cl
28 Apr 2026
Tutorials

Propagation Structure-Semantic Transfer Learning for Robust Fake News Detection

DGX agent

arXiv:2604.23974v1 Announce Type: new Abstract: Fake news generally refers to false information that is spread deliberately to deceive people, which has detrimental social effects. Existing fake news

tutorialsarxiv-cs-cl
28 Apr 2026
Model Releases

Reading in the Dark: Low-light Scene Text Recognition

DGX agent

arXiv:2604.23685v1 Announce Type: new Abstract: Accurate text recognition in low-light environments is essential for intelligent systems in applications ranging from autonomous vehicles to smart surve

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

DGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Scalable LLM-based Coding of Dialogue in Healthcare Simulation: Balancing Coding Performance, Processing Time, and Environmental Impact

DGX agent

arXiv:2604.23255v1 Announce Type: cross Abstract: Research shows that dialogue, the interactive process through which participants articulate their thinking, plays a central role in constructing share

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

DGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Symbolic recovery of PDEs from measurement data

DGX agent

arXiv:2602.15603v2 Announce Type: replace Abstract: Models based on partial differential equations (PDEs) are powerful for describing a wide range of complex phenomena in the natural sciences. Accurat

researcharxiv-cs-lg
28 Apr 2026
Research

Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

DGX agent

arXiv:2604.23964v1 Announce Type: cross Abstract: Patients with dementia typically exhibit cognitive impairment, which is routinely assessed using the Mini-Mental State Examination (MMSE). Concurrentl

researcharxiv-cs-ai
28 Apr 2026
Safety

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

DGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

safetyarxiv-cs-ai
28 Apr 2026
Agents

Think Anywhere in Code Generation

DGX agent

arXiv:2603.29957v3 Announce Type: replace-cross Abstract: Recent advances in reasoning Large Language Models (LLMs) have primarily relied on upfront thinking, where reasoning occurs before final answe

agentsarxiv-cs-lg
28 Apr 2026
Local Ai

Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set

DGX agent

arXiv:2502.03669v3 Announce Type: replace-cross Abstract: AI methods, such as generative models and reinforcement learning, have recently been applied to combinatorial optimization (CO) problems, espe

local-aiarxiv-cs-ai
28 Apr 2026
Agents

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

DGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

DGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

DGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

DGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Adapting MLLMs for Nuanced Video Retrieval

DGX agent

arXiv:2512.13511v2 Announce Type: replace Abstract: Our objective is to build an embedding model that captures the nuanced relationship between a search query and candidate videos. We cover three aspe

researcharxiv-cs-cv
27 Apr 2026
Safety

AgentBound: Securing Execution Boundaries of AI Agents

DGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

DGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

model-releasesarxiv-cs-cv
27 Apr 2026
Research

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

DGX agent

arXiv:2601.14287v2 Announce Type: replace Abstract: External memory systems are pivotal for enabling Large Language Model (LLM) agents to maintain persistent knowledge and perform long-horizon decisio

researcharxiv-cs-lg
27 Apr 2026
Tutorials

CLVAE: A Variational Autoencoder for Long-Term Customer Revenue Forecasting

DGX agent

arXiv:2604.22636v1 Announce Type: cross Abstract: Predicting customers' long-term revenue from sparse and irregular transaction data is central to marketing resource allocation in non-contractual sett

tutorialsarxiv-cs-lg
27 Apr 2026
Safety

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

DGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

safetyarxiv-cs-ai
27 Apr 2026
Research

Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation

DGX agent

arXiv:2510.19592v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) demonstrate strong video understanding by attending to visual tokens relevant to textual queries. To direct

researcharxiv-cs-cv
27 Apr 2026
Model Releases

EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges

DGX agent

arXiv:2604.22595v1 Announce Type: new Abstract: CLIP has demonstrated strong generalization in visual domains through natural language supervision, even for video action recognition. However, most exi

model-releasesarxiv-cs-cv
27 Apr 2026
Safety

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

DGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

safetyarxiv-cs-cv
27 Apr 2026
Research

LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs

DGX agent

arXiv:2604.22050v1 Announce Type: cross Abstract: Transformers are mostly relying on softmax attention, which introduces quadratic complexity with respect to sequence length and remains a major bottle

researcharxiv-cs-cl
27 Apr 2026
Model Releases

LLMs as Assessors: Right for the Right Reason?

DGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

model-releasesarxiv-cs-cl
27 Apr 2026
Research

NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs

DGX agent

arXiv:2511.07003v2 Announce Type: replace Abstract: Large language models have significantly advanced Multilingual Machine Translation (MMT), yet scaling to many languages while keeping quality robust

researcharxiv-cs-cl
27 Apr 2026
Model Releases

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

DGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

model-releasesarxiv-cs-cv
27 Apr 2026
← Previous
1…456457458459460…1119
Next →