AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Fine-tuning DeepSeek-OCR-2 for Molecular Structure Recognition

DGX agent

arXiv:2604.03476v2 Announce Type: replace-cross Abstract: Optical Chemical Structure Recognition (OCSR) is critical for converting 2D molecular diagrams from printed literature into machine-readable f

model-releasesarxiv-cs-ai
22 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LLMs Know They're Wrong and Agree Anyway: The Shared Sycophancy-Lying Circuit

DGX agent

arXiv:2604.19117v1 Announce Type: new Abstract: When a language model agrees with a user's false belief, is it failing to detect the error, or noticing and agreeing anyway? We show the latter. Across

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

LSTM-MAS: A Long Short-Term Memory Inspired Multi-Agent System for Long-Context Understanding

DGX agent

arXiv:2601.11913v2 Announce Type: replace-cross Abstract: Effectively processing long contexts remains a fundamental yet unsolved challenge for large language models (LLMs). Existing single-LLM-based

model-releasesarxiv-cs-ai
22 Apr 2026
Tutorials

MapPFN: Learning Causal Perturbation Maps in Context

DGX agent

arXiv:2601.21092v2 Announce Type: replace Abstract: Planning effective interventions in biological systems requires treatment-effect models that adapt to unseen biological contexts by identifying thei

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Mechanistic Anomaly Detection via Functional Attribution

DGX agent

arXiv:2604.18970v1 Announce Type: new Abstract: We can often verify the correctness of neural network outputs using ground truth labels, but we cannot reliably determine whether the output was produce

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

Multi-Domain Learning with Global Expert Mapping

DGX agent

arXiv:2604.18842v1 Announce Type: new Abstract: Human perception generalizes well across different domains, but most vision models struggle beyond their training data. This gap motivates multi-dataset

model-releasesarxiv-cs-cv
22 Apr 2026
Local Ai

Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?

DGX agent

arXiv:2604.19399v1 Announce Type: new Abstract: Federated learning (FL) is a key paradigm for distributed model learning across decentralized data sources. Communication in each FL round typically con

local-aiarxiv-cs-lg
22 Apr 2026
Model Releases

PriorGuide: Test-Time Prior Adaptation for Simulation-Based Inference

DGX agent

arXiv:2510.13763v2 Announce Type: replace-cross Abstract: Amortized simulator-based inference offers a powerful framework for tackling Bayesian inference in computational fields such as engineering or

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Probing for Reading Times

DGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Safe Continual Reinforcement Learning in Non-stationary Environments

DGX agent

arXiv:2604.19737v1 Announce Type: new Abstract: Reinforcement learning (RL) offers a compelling data-driven paradigm for synthesizing controllers for complex systems when accurate physical models are

model-releasesarxiv-cs-lg
22 Apr 2026
Research

SimDiff: Depth Pruning via Similarity and Difference

DGX agent

arXiv:2604.19520v1 Announce Type: new Abstract: Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standar

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

DGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

DGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

DGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

DGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

DGX agent

arXiv:2604.02368v4 Announce Type: replace Abstract: As Large Language Models (LLMs) exhibit plateauing performance on conventional benchmarks, a pivotal challenge persists: evaluating their proficienc

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules

DGX agent

arXiv:2604.16550v1 Announce Type: new Abstract: Despite the high accuracy of 'black box' deep learning models, drug discovery still relies on protein-ligand interaction principles and heuristics. To i

model-releasesarxiv-cs-lg
21 Apr 2026
Research

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning

DGX agent

arXiv:2510.14738v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have rapidly advanced from perception tasks to complex multi-step reasoning, yet reinforcement learning wit

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

DGX agent

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

model-releasesarxiv-cs-lg
21 Apr 2026
Research

BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs

DGX agent

arXiv:2604.17629v1 Announce Type: new Abstract: Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-c

researcharxiv-cs-cv
21 Apr 2026
Model Releases

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

DGX agent

arXiv:2507.20879v3 Announce Type: replace Abstract: The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

DGX agent

arXiv:2509.20360v3 Announce Type: replace Abstract: Recent advances in foundation models highlight a clear trend toward unification and scaling, showing emergent capabilities across diverse domains. W

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

DGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

'Faithful to What?' On the Limits of Fidelity-Based Explanations

DGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

FedOBP: Federated Optimal Brain Personalization through Cloud-Edge Element-wise Decoupling

DGX agent

arXiv:2604.16574v1 Announce Type: new Abstract: Federated Learning (FL) faces challenges from client data heterogeneity and resource-constrained mobile devices, which can degrade model accuracy. Perso

model-releasesarxiv-cs-lg
21 Apr 2026
Tutorials

From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation

DGX agent

arXiv:2604.17455v1 Announce Type: new Abstract: Visual prompting has emerged as a powerful method for adapting pre-trained models to new domains without updating model parameters. However, existing pr

tutorialsarxiv-cs-cv
21 Apr 2026
Model Releases

From Inheritance to Saturation: Disentangling the Evolution of Visual Redundancy for Architecture-Aware MLLM Inference Acceleration

DGX agent

arXiv:2604.16462v1 Announce Type: new Abstract: High-resolution Multimodal Large Language Models (MLLMs) face prohibitive computational costs during inference due to the explosion of visual tokens. Ex

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

DGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Guardrails in Logit Space: Safety Token Regularization for LLM Alignment

DGX agent

arXiv:2604.17210v1 Announce Type: new Abstract: Fine-tuning well-aligned large language models (LLMs) on new domains often degrades their safety alignment, even when using benign datasets. Existing sa

model-releasesarxiv-cs-lg
21 Apr 2026
Research

How Much Data is Enough? The Zeta Law of Discoverability in Biomedical Data, featuring the enigmatic Riemann zeta function

DGX agent

arXiv:2604.17581v1 Announce Type: new Abstract: How much data is enough to make a scientific discovery? As biomedical datasets scale to millions of samples and AI models grow in capacity, progress inc

researcharxiv-cs-lg
21 Apr 2026
Model Releases

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

DGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Live Avatar: Streaming Real-time Audio-Driven Avatar Generation with Infinite Length

DGX agent

arXiv:2512.04677v5 Announce Type: replace Abstract: Audio-driven avatar interaction demands real-time, streaming, and infinite-length generation -- capabilities fundamentally at odds with the sequenti

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

DGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems

DGX agent

arXiv:2503.16549v2 Announce Type: replace Abstract: Despite strong results on many tasks, multimodal large language models (MLLMs) still underperform on visual mathematical problem solving, especially

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Maximizing Local Entropy Where It Matters: Prefix-Aware Localized LLM Unlearning

DGX agent

arXiv:2601.03190v3 Announce Type: replace Abstract: Machine unlearning aims to forget sensitive knowledge from Large Language Models (LLMs) while maintaining general utility. However, existing approac

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

MedProbeBench: Systematic Benchmarking at Deep Evidence Integration for Expert-level Medical Guideline

DGX agent

arXiv:2604.18418v1 Announce Type: new Abstract: Recent advances in deep research systems enable large language models to retrieve, synthesize, and reason over large-scale external knowledge. In medici

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Mitigating Multimodal Hallucination via Phase-wise Self-reward

DGX agent

arXiv:2604.17982v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) still struggle with vision hallucination, where generated responses are inconsistent with the visual input. Exist

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Multiplication in Multimodal LLMs: Computation with Text, Image, and Audio Inputs

DGX agent

arXiv:2604.18203v1 Announce Type: new Abstract: Multimodal LLMs can accurately perceive numerical content across modalities yet fail to perform exact multi-digit multiplication when the identical unde

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

DGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

DGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

DGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

QU-NLP at QIAS 2026: Multi-Stage QLoRA Fine-Tuning for Arabic Islamic Inheritance Reasoning

DGX agent

arXiv:2604.16396v1 Announce Type: new Abstract: Islamic inheritance law (ilm al-mawar{i}th) presents a challenging domain for evaluating large language models' structured reasoning capabilities, requi

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Sessa: Selective State Space Attention

DGX agent

arXiv:2604.18580v1 Announce Type: cross Abstract: Modern sequence models are dominated by Transformers, where self-attention mixes information from the visible context in an input-dependent way. Howev

researcharxiv-cs-cl
21 Apr 2026
Model Releases

The Cognitive Penalty: Ablating System 1 and System 2 Reasoning in Edge-Native SLMs for Decentralized Consensus

DGX agent

arXiv:2604.16913v1 Announce Type: cross Abstract: Decentralized Autonomous Organizations (DAOs) are inclined explore Small Language Models (SLMs) as edge-native constitutional firewalls to vet proposa

model-releasesarxiv-cs-cl
21 Apr 2026
Research

ThinkBrake: Efficient Reasoning via Log-Probability Margin Guided Decoding

DGX agent

arXiv:2510.00546v5 Announce Type: replace Abstract: Large Reasoning Models (LRMs) allocate substantial inference-time compute to Chain-of-Thought (CoT) reasoning, improving performance on mathematics,

researcharxiv-cs-cl
21 Apr 2026
Model Releases

TimeColor: Flexible Reference Colorization via Temporal Concatenation

DGX agent

arXiv:2601.00296v2 Announce Type: replace Abstract: Most colorization models condition only on a single reference, typically the first frame of the scene. However, this approach ignores other sources

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

TinySR: Pruning Diffusion for Real-World Image Super-Resolution

DGX agent

arXiv:2508.17434v2 Announce Type: replace Abstract: Real-world image super-resolution (Real-ISR) focuses on recovering high-quality images from low-resolution inputs that suffer from complex degradati

model-releasesarxiv-cs-cv
21 Apr 2026
Research

TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

DGX agent

arXiv:2603.19684v2 Announce Type: replace Abstract: Automatic tooth segmentation and identification from intra-oral scanned 3D models are fundamental problems in digital dentistry, yet most existing a

researcharxiv-cs-cv
21 Apr 2026
← Previous
1…330331332333334…1065
Next →