AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,154 results
Model Releases

Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining

DGX agent

arXiv:2604.16391v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have shown great potential in building generalist robots, but still face a dilemma-misalignment of 2D image foreca

model-releasesarxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

eCP: Equivariant Conformal Prediction with pre-trained models

DGX agent

arXiv:2602.03986v2 Announce Type: replace Abstract: Conformal prediction, a post-hoc, distribution-free, finite-sample method of uncertainty quantification that offers formal coverage guarantees under

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Emergent Misalignment via In-Context Learning: Narrow in-context examples can produce broadly misaligned LLMs

DGX agent

arXiv:2510.11288v4 Announce Type: replace Abstract: Recent work has shown that narrow finetuning can produce broadly misaligned LLMs, a phenomenon termed emergent misalignment (EM). While concerning,

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Forecasting Ionospheric Irregularities on GNSS Lines of Sight Using Dynamic Graphs with Ephemeris Conditioning

DGX agent

arXiv:2604.18379v1 Announce Type: new Abstract: Most data-driven ionospheric forecasting models operate on gridded products, which do not preserve the time-varying sampling structure of satellite-base

model-releasesarxiv-cs-lg
21 Apr 2026
Applications

FUSE: Ensembling Verifiers with Zero Labeled Data

DGX agent

arXiv:2604.18547v1 Announce Type: cross Abstract: Verification of model outputs is rapidly emerging as a key primitive for both training and real-world deployment of large language models (LLMs). In p

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

HiP-LoRA: Budgeted Spectral Plasticity for Robust Low-Rank Adaptation

DGX agent

arXiv:2604.17751v1 Announce Type: cross Abstract: Adapting foundation models under resource budgets relies heavily on Parameter-Efficient Fine-Tuning (PEFT), with LoRA being a standard modular solutio

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Kimi Kimi 。 Kimi Kimi Kimi Kimi Kimi Kimi ollama run kimi-k2.6:cloud

DGX agent

This appears to be a social media post from Ollama's X account regarding a model run command for 'kimi-k2.6:cloud,' likely announcing or demonstrating how to execute this specific AI model variant usi

local-aiollama--x
21 Apr 2026
Applications

Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes

DGX agent

arXiv:2604.18381v1 Announce Type: cross Abstract: Fine-tuning Large Language Models (LLMs) typically relies on large quantities of high-quality annotated data, or questions with well-defined ground tr

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

LoRA on the Go: Instance-level Dynamic LoRA Selection and Merging

DGX agent

arXiv:2511.07129v3 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) has emerged as a parameter-efficient approach for fine-tuning large language models. However, conventional LoRA adapters

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Medical thinking with multiple images

DGX agent

arXiv:2604.16506v1 Announce Type: cross Abstract: Large language models perform well on many medical QA benchmarks, but real clinical reasoning often requires integrating evidence across multiple imag

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Missing Pattern Tree based Decision Grouping and Ensemble for Enhancing Pair Utilization in Deep Incomplete Multi-View Clustering

DGX agent

arXiv:2512.21510v2 Announce Type: replace-cross Abstract: Real-world multi-view data often exhibit highly inconsistent missing patterns, posing significant challenges for incomplete multi-view cluster

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

On the Importance and Evaluation of Narrativity in Natural Language AI Explanations

DGX agent

arXiv:2604.18311v1 Announce Type: new Abstract: Explainable AI (XAI) aims to make the behaviour of machine learning models interpretable, yet many explanation methods remain difficult to understand. T

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

PAC-Bayes Bounds for Gibbs Posteriors via Singular Learning Theory

DGX agent

arXiv:2604.17219v1 Announce Type: cross Abstract: We derive explicit non-asymptotic PAC-Bayes generalization bounds for Gibbs posteriors, that is, data-dependent distributions over model parameters ob

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

DGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

model-releasesarxiv-cs-cv
21 Apr 2026
Research

PCM-NeRF: Probabilistic Camera Modeling for Neural Radiance Fields under Pose Uncertainty

DGX agent

arXiv:2604.17831v1 Announce Type: new Abstract: Neural surface reconstruction methods typically treat camera poses as fixed values, assuming perfect accuracy from Structure-from-Motion (SfM) systems.

researcharxiv-cs-cv
21 Apr 2026
Model Releases

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning

DGX agent

arXiv:2604.17800v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have gained much attention from the research community thanks to their strength in translating multimodal observat

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design

DGX agent

arXiv:2604.17175v1 Announce Type: new Abstract: We introduce RosettaSearch, an inference-time multi-objective optimization approach for protein sequence optimization. We use large language models (LLM

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

SetFlow: Generating Structured Sets of Representations for Multiple Instance Learning

DGX agent

arXiv:2604.16362v1 Announce Type: cross Abstract: Data scarcity and weak supervision continue to limit the performance of machine learning models in many real-world applications, such as mammography,

model-releasesarxiv-cs-cv
21 Apr 2026
Research

Tool Learning Needs Nothing More Than a Free 8B Language Model

DGX agent

arXiv:2604.17739v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a prevalent paradigm for training tool calling agents, which typically requires online interactive environments

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Universally Empowering Zeroth-Order Optimization via Adaptive Layer-wise Sampling

DGX agent

arXiv:2604.18264v1 Announce Type: new Abstract: Zeroth-Order optimization presents a promising memory-efficient paradigm for fine-tuning Large Language Models by relying solely on forward passes. Howe

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Beyond Distribution Sharpening: The Importance of Task Rewards

DGX agent

arXiv:2604.16259v1 Announce Type: cross Abstract: Frontier models have demonstrated exceptional capabilities following the integration of task-reward-based reinforcement learning (RL) into their train

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond MCQ: An Open-Ended Arabic Cultural QA Benchmark with Dialect Variants

DGX agent

arXiv:2510.24328v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used to answer everyday questions, yet their performance on culturally grounded and dialectal co

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Brain Score Tracks Shared Properties of Languages: Evidence from Many Natural Languages and Structured Sequences

DGX agent

arXiv:2604.15503v1 Announce Type: new Abstract: Recent breakthroughs in language models (LMs) using neural networks have raised the question: how similar are these models' processing to human language

researcharxiv-cs-cl
20 Apr 2026
Model Releases

Frequency-Aware Flow Matching for High-Quality Image Generation

DGX agent

arXiv:2604.15521v1 Announce Type: new Abstract: Flow matching models have emerged as a powerful framework for realistic image generation by learning to reverse a corruption process that progressively

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Long-Term Memory for VLA-based Agents in Open-World Task Execution

DGX agent

arXiv:2604.15671v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated significant potential for embodied decision-making; however, their application in complex chemical

safetyarxiv-cs-ro
20 Apr 2026
Safety

On the Rejection Criterion for Proxy-based Test-time Alignment

DGX agent

arXiv:2604.16146v1 Announce Type: new Abstract: Recent works proposed test-time alignment methods that rely on a small aligned model as a proxy that guides the generation of a larger base (unaligned)

safetyarxiv-cs-cl
20 Apr 2026
Applications

ProtoTTA: Prototype-Guided Test-Time Adaptation

DGX agent

arXiv:2604.15494v1 Announce Type: cross Abstract: Deep networks that rely on prototypes-interpretable representations that can be related to the model input-have gained significant attention for balan

applicationsarxiv-cs-cv
20 Apr 2026
Model Releases

RefereeBench: Are Video MLLMs Ready to be Multi-Sport Referees

DGX agent

arXiv:2604.15736v1 Announce Type: cross Abstract: While Multimodal Large Language Models (MLLMs) excel at generic video understanding, their ability to support specialized, rule-grounded decision-maki

model-releasesarxiv-cs-cl
20 Apr 2026
Agents

VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation

DGX agent

arXiv:2510.27617v2 Announce Type: replace Abstract: Automation of Register Transfer Level (RTL) design can help developers meet increasing computational demands. Large Language Models (LLMs) show prom

agentsarxiv-cs-ai
20 Apr 2026
Tutorials

Attention to Mamba: A Recipe for Cross-Architecture Distillation

DGX agent

arXiv:2604.14191v1 Announce Type: new Abstract: State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher thro

tutorialsarxiv-cs-cl
17 Apr 2026
Safety

Continuous-time reinforcement learning: ellipticity enables model-free value function approximation

DGX agent

arXiv:2602.06930v2 Announce Type: replace Abstract: We study off-policy reinforcement learning for controlling continuous-time Markov diffusion processes with discrete-time observations and actions. W

safetyarxiv-cs-lg
17 Apr 2026
Tutorials

KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality

DGX agent

arXiv:2506.19807v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs), particularly slow-thinking models, often exhibit severe hallucination, outputting incorrect content due to an in

tutorialsarxiv-cs-cl
17 Apr 2026
Model Releases

LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learning

DGX agent

arXiv:2604.14922v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a critical driver for enhancing the reasoning capabilities of Large Language Models (LLMs). While recent ad

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesis

DGX agent

arXiv:2604.15093v1 Announce Type: cross Abstract: Mobile agents powered by vision-language models have demonstrated impressive capabilities in automating mobile tasks, with recent leading models achie

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Purging the Gray Zone: Latent-Geometric Denoising for Precise Knowledge Boundary Awareness

DGX agent

arXiv:2604.14324v1 Announce Type: new Abstract: Large language models (LLMs) often exhibit hallucinations due to their inability to accurately perceive their own knowledge boundaries. Existing abstent

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Quantum-inspired tensor networks in machine learning models

DGX agent

arXiv:2604.14287v1 Announce Type: new Abstract: Tensor networks were developed in the context of many-body physics as compressed representations of multiparticle quantum states. These representations

researcharxiv-cs-lg
17 Apr 2026
Safety

SeaAlert: Critical Information Extraction From Maritime Distress Communications with Large Language Models

DGX agent

arXiv:2604.14163v1 Announce Type: new Abstract: Maritime distress communications transmitted over very high frequency (VHF) radio are safety-critical voice messages used to report emergencies at sea.

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Secure and Privacy-Preserving Vertical Federated Learning

DGX agent

arXiv:2604.13474v1 Announce Type: cross Abstract: We propose a novel end-to-end privacy-preserving framework, instantiated by three efficient protocols for different deployment scenarios, covering bot

model-releasesarxiv-cs-ai
17 Apr 2026
Safety

Step-level Denoising-time Diffusion Alignment with Multiple Objectives

DGX agent

arXiv:2604.14379v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as a powerful tool for aligning diffusion models with human preferences, typically by optimizing a single rewa

safetyarxiv-cs-cv
17 Apr 2026
Safety

StoryCoder: Narrative Reformulation for Structured Reasoning in LLM Code Generation

DGX agent

arXiv:2604.14631v1 Announce Type: new Abstract: Effective code generation requires both model capability and a problem representation that carefully structures how models reason and plan. Existing app

safetyarxiv-cs-cl
17 Apr 2026
Research

Threshold Differential Attention for Sink-Free, Ultra-Sparse, and Non-Dispersive Language Modeling

DGX agent

arXiv:2601.12145v2 Announce Type: replace Abstract: Softmax attention struggles with long contexts due to structural limitations: the strict sum-to-one constraint forces attention sinks on irrelevant

researcharxiv-cs-lg
17 Apr 2026
Model Releases

A Study of Failure Modes in Two-Stage Human-Object Interaction Detection

DGX agent

arXiv:2604.13448v1 Announce Type: new Abstract: Human-object interaction (HOI) detection aims to detect interactions between humans and objects in images. While recent advances have improved performan

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Asymmetric-Loss-Guided Hybrid CNN-BiLSTM-Attention Model for Industrial RUL Prediction with Interpretable Failure Heatmaps

DGX agent

arXiv:2604.13459v1 Announce Type: new Abstract: Turbofan engine degradation under sustained operational stress necessitates robust prognostic systems capable of accurately estimating the Remaining Use

safetyarxiv-cs-lg
16 Apr 2026
Safety

Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration

DGX agent

arXiv:2604.13705v1 Announce Type: new Abstract: Fairness in language models is typically studied as a property of a single, centrally optimized model. As large language models become increasingly agen

safetyarxiv-cs-cl
16 Apr 2026
Safety

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

DGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

safetyarxiv-cs-ro
16 Apr 2026
Research

Dual-Enhancement Product Bundling: Bridging Interactive Graph and Large Language Model

DGX agent

arXiv:2604.14030v1 Announce Type: new Abstract: Product bundling boosts e-commerce revenue by recommending complementary item combinations. However, existing methods face two critical challenges: (1)

researcharxiv-cs-cl
16 Apr 2026
Model Releases

From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs

DGX agent

arXiv:2604.14137v1 Announce Type: new Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'':

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

Hierarchical DLO Routing with Reinforcement Learning and In-Context Vision-language Models

DGX agent

arXiv:2510.19268v2 Announce Type: replace-cross Abstract: Long-horizon routing tasks of deformable linear objects (DLOs), such as cables and ropes, are common in industrial assembly lines and everyday

agentsarxiv-cs-lg
16 Apr 2026
← Previous
1…345346347348349…1337
Next →