AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Applications

UAF: A Unified Audio Front-end LLM for Full-Duplex Speech Interaction

DGX agent

arXiv:2604.19221v1 Announce Type: new Abstract: Full-duplex speech interaction, as the most natural and intuitive mode of human communication, is driving artificial intelligence toward more human-like

applicationsarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

When Safety Fails Before the Answer: Benchmarking Harmful Behavior Detection in Reasoning Chains

DGX agent

arXiv:2604.19001v1 Announce Type: new Abstract: Large reasoning models (LRMs) produce complex, multi-step reasoning traces, yet safety evaluation remains focused on final outputs, overlooking how harm

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Who Shapes Brazil's Vaccine Debate? Semi-Supervised Modeling of Stance and Polarization in YouTube's Media Ecosystem

DGX agent

arXiv:2604.18586v1 Announce Type: cross Abstract: Vaccination remains a cornerstone of global public health, yet the COVID-19 pandemic exposed how online misinformation, political polarization, and de

researcharxiv-cs-ai
22 Apr 2026
Research

A Mechanism Study of Delayed Loss Spikes in Batch-Normalized Linear Models

DGX agent

arXiv:2604.16809v1 Announce Type: cross Abstract: Delayed loss spikes have been reported in neural-network training, but existing theory mainly explains earlier non-monotone behavior caused by overly

researcharxiv-cs-lg
21 Apr 2026
Research

A Model and Estimation of the Bitcoin Transaction Fee

DGX agent

arXiv:2604.17183v1 Announce Type: cross Abstract: Bitcoin transaction fees will become more important as the block subsidy declines, but fee formation is hard to study with blockchain data alone becau

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL

DGX agent

arXiv:2604.17073v1 Announce Type: new Abstract: Reinforcement fine-tuning improves the reasoning ability of large language models, but it can also encourage them to answer unanswerable queries by gues

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Adaptive Forensic Feature Refinement via Intrinsic Importance Perception

DGX agent

arXiv:2604.16879v1 Announce Type: new Abstract: With the rapid development of generative models and multimodal content editing technologies, the key challenge faced by synthetic image detection (SID)

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Are We Using the Right Benchmark: An Evaluation Framework for Visual Token Compression Methods

DGX agent

arXiv:2510.07143v3 Announce Type: replace Abstract: Recent efforts to accelerate inference in Multimodal Large Language Models (MLLMs) have largely focused on visual token compression. The effectivene

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Beyond Reproduction: A Paired-Task Framework for Assessing LLM Comprehension and Creativity in Literary Translation

DGX agent

arXiv:2604.18169v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for creative tasks such as literary translation. Yet translational creativity remains underexplored a

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Cat-DPO: Category-Adaptive Safety Alignment

DGX agent

arXiv:2604.17299v1 Announce Type: new Abstract: Aligning large language models with human preferences must balance two competing goals: responding helpfully to legitimate requests and reliably refusin

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

CoLLM: A Unified Framework for Co-execution of LLMs Federated Fine-tuning and Inference

DGX agent

arXiv:2604.16400v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly adopted in edge intelligence to power domain-specific applications and personalized services, the qua

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

DGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

model-releasesarxiv-cs-cv
21 Apr 2026
Research

eCP: Equivariant Conformal Prediction with pre-trained models

DGX agent

arXiv:2602.03986v2 Announce Type: replace Abstract: Conformal prediction, a post-hoc, distribution-free, finite-sample method of uncertainty quantification that offers formal coverage guarantees under

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

DGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Forecasting Ionospheric Irregularities on GNSS Lines of Sight Using Dynamic Graphs with Ephemeris Conditioning

DGX agent

arXiv:2604.18379v1 Announce Type: new Abstract: Most data-driven ionospheric forecasting models operate on gridded products, which do not preserve the time-varying sampling structure of satellite-base

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

FUSE: Ensembling Verifiers with Zero Labeled Data

DGX agent

arXiv:2604.18547v1 Announce Type: cross Abstract: Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In p

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

DGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

model-releasesarxiv-cs-cl
21 Apr 2026
Applications

Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes

DGX agent

arXiv:2604.18381v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) typically relies on large quantities of high-quality annotated data, or questions with well-defined ground tr

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations

DGX agent

arXiv:2604.18311v1 Announce Type: new Abstract: Explainable AI (XAI) aims to make the behaviour of machine learning models interpretable, yet many explanation methods remain difficult to understand. T

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

DGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

DGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

model-releasesarxiv-cs-cv
21 Apr 2026
Research

PCM-NeRF: Probabilistic Camera Modeling for Neural Radiance Fields under Pose Uncertainty

DGX agent

arXiv:2604.17831v1 Announce Type: new Abstract: Neural surface reconstruction methods typically treat camera poses as fixed values, assuming perfect accuracy from Structure-from-Motion (SfM) systems.

researcharxiv-cs-cv
21 Apr 2026
Model Releases

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning

DGX agent

arXiv:2604.17800v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observat

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design

DGX agent

arXiv:2604.17175v1 Announce Type: new Abstract: We introduce RosettaSearch, an inference-time multi-objective optimization approach for protein sequence optimization. We use large language models (LLM

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

SetFlow: Generating Structured Sets of Representations for Multiple Instance Learning

DGX agent

arXiv:2604.16362v1 Announce Type: cross Abstract: Data scarcity and weak supervision continue to limit the performance of machine learning models in many real-world applications, such as mammography,

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Tool Learning Needs Nothing More Than a Free 8B Language Model

DGX agent

arXiv:2604.17739v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a prevalent paradigm for training tool calling agents, which typically requires online interactive environments

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling

DGX agent

arXiv:2604.18264v1 Announce Type: new Abstract: Zeroth-Order optimization presents a promising memory-efficient paradigm for fine-tuning Large Language Models by relying solely on forward passes. Howe

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Beyond Distribution Sharpening: The Importance of Task Rewards

DGX agent

arXiv:2604.16259v1 Announce Type: cross Abstract: Frontier models have demonstrated exceptional capabilities following the integration of task-reward-based reinforcement learning (RL) into their train

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants

DGX agent

arXiv:2510.24328v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used to answer everyday questions, yet their performance on culturally grounded and dialectal co

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

DGX agent

arXiv:2604.15503v1 Announce Type: new Abstract: Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Frequency-Aware Flow Matching for High-Quality Image Generation

DGX agent

arXiv:2604.15521v1 Announce Type: new Abstract: Flow matching models have emerged as a powerful framework for realistic image generation by learning to reverse a corruption process that progressively

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Safety

On the Rejection Criterion for Proxy-based Test-time Alignment

DGX agent

arXiv:2604.16146v1 Announce Type: new Abstract: Recent works proposed test-time alignment methods that rely on a small aligned model as a proxy that guides the generation of a larger base (unaligned)

safetyarxiv-cs-cl
20 Apr 2026
Applications

ProtoTTA: Prototype-Guided Test-Time Adaptation

DGX agent

arXiv:2604.15494v1 Announce Type: cross Abstract: Deep networks that rely on prototypes-interpretable representations that can be related to the model input-have gained significant attention for balan

applicationsarxiv-cs-cv
20 Apr 2026
Model Releases

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees

DGX agent

arXiv:2604.15736v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) excel at generic video understanding, their ability to support specialized, rule-grounded decision-maki

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

DGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

agentsarxiv-cs-ai
20 Apr 2026
Tutorials

Attention to Mamba: A Recipe for Cross-Architecture Distillation

DGX agent

arXiv:2604.14191v1 Announce Type: new Abstract: State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher thro

tutorialsarxiv-cs-cl
17 Apr 2026
Safety

Continuous-time reinforcement learning: ellipticity enables model-free value function approximation

DGX agent

arXiv:2602.06930v2 Announce Type: replace Abstract: We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. W

safetyarxiv-cs-lg
17 Apr 2026
Tutorials

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality

DGX agent

arXiv:2506.19807v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly slow-thinking models, often exhibit severe hallucination, outputting incorrect content due to an in

tutorialsarxiv-cs-cl
17 Apr 2026
Model Releases

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

DGX agent

arXiv:2604.14922v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a critical driver for enhancing the reasoning capabilities of Large Language Models (LLMs). While recent ad

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis

DGX agent

arXiv:2604.15093v1 Announce Type: cross Abstract: Mobile agents powered by vision-language models have demonstrated impressive capabilities in automating mobile tasks, with recent leading models achie

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness

DGX agent

arXiv:2604.14324v1 Announce Type: new Abstract: Large language models (LLMs) often exhibit hallucinations due to their inability to accurately perceive their own knowledge boundaries. Existing abstent

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Quantum-inspired tensor networks in machine learning models

DGX agent

arXiv:2604.14287v1 Announce Type: new Abstract: Tensor networks were developed in the context of many-body physics as compressed representations of multiparticle quantum states. These representations

researcharxiv-cs-lg
17 Apr 2026
Safety

SeaAlert: Critical Information Extraction From Maritime Distress Communications with Large Language Models

DGX agent

arXiv:2604.14163v1 Announce Type: new Abstract: Maritime distress communications transmitted over very high frequency (VHF) radio are safety-critical voice messages used to report emergencies at sea.

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Secure and Privacy-Preserving Vertical Federated Learning

DGX agent

arXiv:2604.13474v1 Announce Type: cross Abstract: We propose a novel end-to-end privacy-preserving framework, instantiated by three efficient protocols for different deployment scenarios, covering bot

model-releasesarxiv-cs-ai
17 Apr 2026
← Previous
1…270271272273274…1058
Next →