AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,509 results
3 Jun 2026

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation

Model ReleasesDGX agent

arXiv:2606.03784v1 Announce Type: new Abstract: Embodied chain-of-thought (CoT) aims to bridge linguistic reasoning and robotic control, but its effective form and integration strategy remain underexp

Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta Solvers

ResearchDGX agent

arXiv:2502.08834v4 Announce Type: replace-cross Abstract: Deep generative models based on neural differential equations have become state-of-the-art for many generation tasks. These models rely on ODE

scTranslation: A Comprehensive Benchmark for Single-Cell Multi-Omics Modality Translation

Model ReleasesDGX agent

arXiv:2606.03906v1 Announce Type: new Abstract: Simultaneous measurement of multiple omics modalities in single cells enables researchers to gain a more comprehensive understanding of cellular states

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

SafetyDGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning

Model ReleasesDGX agent

arXiv:2606.03503v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have achieved remarkable progress thanks to Reinforcement Learning with Verifiable Rewards (RLVR) on Chain-of-Thoughts (Co

Unified Video-Action Joint Denoising for Dexterous Action and Data Generation

SafetyDGX agent

arXiv:2606.03868v1 Announce Type: new Abstract: Recent world action models leverage video foundation models by aligning broad visual-dynamics priors with executable robot actions. We revisit this alig

2 Jun 2026

AblationBench: Evaluating Automated Planning of Ablations in Empirical AI Research

Model ReleasesDGX agent

arXiv:2507.08038v3 Announce Type: replace-cross Abstract: Language model agents are increasingly used to automate scientific research, yet evaluating their scientific contributions remains a challenge

Announcing Spanner Graph algorithms: Google-grade intelligence for connected data

Model ReleasesDGX agent

At Google Cloud Next, we announced the preview of graph algorithms with Spanner Graph, bringing Google Research’s state-of-the-art graph mining capabilities natively to your database. These graph inte

ArrythML: An Autoencoder-Based TinyML Approach for On-Device Arrhythmia Detection on Resource-Constrained Embedded Systems

Local AiDGX agent

arXiv:2606.02256v1 Announce Type: new Abstract: Our work presents a method for ECG segmentation and arrhythmia detection using Tiny Machine Learning (TinyML) models for real-time, on-device inference

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations

ResearchDGX agent

arXiv:2511.20295v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) are minimal and semantically meaningful modifications of the input of a model that alter the model predictions. T

Before and After Temperature: A Distributional View of Creative LLM Generation

Model ReleasesDGX agent

arXiv:2606.01451v1 Announce Type: new Abstract: Reference-free evaluation of large language model (LLM) creativity relies on perplexity, entropy, and top-1 margin. We show that a much stronger signal

Benchmark Dataset for Catalysis on 2D MXenes

Model ReleasesDGX agent

arXiv:2606.00794v1 Announce Type: cross Abstract: Merging first-principles calculations with machine learning (ML), we aim to accelerate the exploration of catalytic behaviour in novel materials. We f

Benchmarking LLM-as-a-Judge for Long-Form Output Evaluation

Model ReleasesDGX agent

arXiv:2606.01629v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly used for long-form generation, reliably evaluating long-form outputs has become a critical challenge. L

Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardware

Model ReleasesDGX agent

arXiv:2606.01338v1 Announce Type: new Abstract: Biopharmaceutical manufacturing organizations operate under regulatory frameworks such as FDA guidance, EU Good Manufacturing Practice (GMP), and the EU

Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA

Model ReleasesDGX agent

arXiv:2602.09492v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is a standard approach for fine-tuning large language models, yet its many variants report conflicting empirical ga

Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification

ResearchDGX agent

arXiv:2510.24078v2 Announce Type: replace Abstract: Text-to-image (T2I) models are increasingly used for synthetic dataset generation, but generating effective synthetic training data for classificati

Beyond Semantic Understanding: Preserving Collaborative Frequency Components in LLM-based Recommendation

Model ReleasesDGX agent

arXiv:2508.10312v2 Announce Type: replace Abstract: Recommender systems in concert with Large Language Models (LLMs) present promising avenues for generating semantically-informed recommendations. How

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention

Model ReleasesDGX agent

arXiv:2512.10414v2 Announce Type: replace Abstract: Recently, reinforcement learning (RL) has become a common choice in enhancing the reasoning capabilities of vision-language models (VLMs). Consideri

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

Model ReleasesDGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

Child-directed speech facilitates production, not comprehension, in BabyLMs

Model ReleasesDGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

Model ReleasesDGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

Model ReleasesDGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

ClinTutor-R1: Advancing Scalable and Robust One-to-Many Alignment in Clinical Socratic Education

SafetyDGX agent

arXiv:2512.05671v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have achieved remarkable success in dyadic (one-on-one) instruction, they face significant challenges in One-to-M

Consistency evaluation of benchmarks used for causal discovery

Model ReleasesDGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

Model ReleasesDGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

Disentanglement-Based Equivariant Learning for Compositional VQA

Model ReleasesDGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

Model ReleasesDGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding

Model ReleasesDGX agent

arXiv:2603.03312v3 Announce Type: replace-cross Abstract: Decoding natural language from non-invasive EEG signals is a promising yet challenging task. However, current state-of-the-art models remain c

Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression

Model ReleasesDGX agent

arXiv:2606.01608v1 Announce Type: new Abstract: Most existing extreme compression methods fail to achieve an optimal rate-distortion-perception trade-off, as they typically prioritize perceptual fidel

GeistBERT: Breathing Life into German NLP

ResearchDGX agent

arXiv:2506.11903v5 Announce Type: replace Abstract: Advances in transformer-based language models have highlighted the benefits of language-specific pre-training on high-quality corpora. In this conte

GraspGen-X: Cross-Embodiment 6-DOF Diffusion-based Grasping

ApplicationsDGX agent

arXiv:2606.00998v1 Announce Type: new Abstract: We study cross-embodiment 6-DOF robot grasping. Unlike prior works, we require the model not only to generalize to novel objects / scenes but also to no

HalleluBERT: Let Every Token That Has Meaning Bear Its Weight

Model ReleasesDGX agent

arXiv:2510.21372v2 Announce Type: replace Abstract: Transformer-based models have advanced NLP, yet Hebrew still lacks a RoBERTa encoder that is trained at scale and released in both base and large va

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/r…

Model ReleasesDGX agent

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/registration If you're looking for a model to use to tackle t

Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations

Model ReleasesDGX agent

arXiv:2606.01204v1 Announce Type: cross Abstract: We investigate whether large language models produce different medical triage recommendations for identical symptoms based solely on the language of t

IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages

Model ReleasesDGX agent

arXiv:2606.01260v1 Announce Type: cross Abstract: Despite being home to more than 1300 ethnic groups and 700 indigenous languages, bias in Large Language Models has not been fully studied in Indonesia

IntraShuffler: A Privacy Preserving Framework for Heterogeneous DP Federated Learning

Model ReleasesDGX agent

arXiv:2606.02563v1 Announce Type: new Abstract: Heterogeneous Differential Privacy (HDP) in Federated Learning (FL) allows clients to select individual privacy budgets (arepsilon_i) according to insti

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

SafetyDGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

Leaf Spectral Reflectance Prediction Using Multi-Head Attention Neural Networks

ResearchDGX agent

arXiv:2606.01432v1 Announce Type: new Abstract: Accurate modeling of leaf spectral reflectance from physiological and biochemical traits is essential for advancing remote sensing applications in plant

Linguistics-Aware Non-Distortionary LLM Watermarking

ResearchDGX agent

arXiv:2606.00613v1 Announce Type: cross Abstract: Watermarking should identify language-model output without degrading quality or limiting verification to the model provider. Multilingual deployment m

Low-Resource Safety Failures Are Action Failures, Not Representation Failures

Model ReleasesDGX agent

arXiv:2606.01196v1 Announce Type: cross Abstract: Safety alignment learned in high-resource languages transfers poorly to low-resource languages. Models refuse harmful prompts in English but fail to r

Make Your VLA More Robust Without More Data By Interleaving Motion Planning

Model ReleasesDGX agent

arXiv:2606.00985v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown remarkable progress for mobile manipulation, but their performance on long-horizon tasks remains poor. Th

MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution

Model ReleasesDGX agent

arXiv:2606.00017v1 Announce Type: new Abstract: Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events t

Navigating the Reality Gap: On-Device Continual Adaptation of ASR for Clinical Telephony

Model ReleasesDGX agent

arXiv:2512.16401v5 Announce Type: replace Abstract: Automatic Speech Recognition (ASR) can significantly reduce documentation burden in clinical workflows, but standard models degrade sharply in real-

On the Generalization in Topology Optimization via Sensitivity-Conditioned Bernoulli Flow Matching

Model ReleasesDGX agent

arXiv:2606.02179v1 Announce Type: cross Abstract: Surrogate models for topology optimization (TO) exhibit highly variable out-of-distribution (OOD) generalization under distribution shifts such as cha

On the Limits of Token Reduction for Efficient Unified Vision Language Training

Model ReleasesDGX agent

arXiv:2606.01503v1 Announce Type: cross Abstract: Unified vision-language models (VLMs) integrate visual understanding and visual generation within a single autoregressive backbone, but their joint tr

On the Theoretical Limitations of Embedding-based Link Prediction

Model ReleasesDGX agent

arXiv:2506.22271v3 Announce Type: replace Abstract: Neural networks often map low-dimensional embeddings to high-dimensional output spaces. Usually, the output layer is linear, which can create a 'ran

OpenDPR: Open-Vocabulary Change Detection via Vision-Centric Diffusion-Guided Prototype Retrieval for Remote Sensing Imagery

Model ReleasesDGX agent

arXiv:2603.27645v2 Announce Type: replace Abstract: Open-vocabulary change detection (OVCD) seeks to recognize arbitrary changes of interest by enabling generalization beyond a fixed set of predefined

Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization

Model ReleasesDGX agent

arXiv:2606.02178v1 Announce Type: cross Abstract: Recent advancements in generative AI have led to image editing models capable of producing realistic forgeries that evade traditional image forgery lo

ProbMoE: Differentiable Probabilistic Routing for Mixture-of-Experts

ResearchDGX agent

arXiv:2606.01509v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models scale by activating only a small subset of experts per token. However, training such models remains challenging becaus

RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting

SafetyDGX agent

arXiv:2606.00147v1 Announce Type: cross Abstract: Domain-specific supervised fine-tuning (SFT) often improves in-domain performance at the cost of degrading a model's general capabilities. We view thi

Realistic noise synthesis reduces bias and improves tissue microstructure estimation with supervised machine learning

Model ReleasesDGX agent

arXiv:2606.02044v1 Announce Type: new Abstract: Diffusion MRI enables non-invasive probing of tissue microstructure, but accurate parameter estimation is challenged by noise-related effects. In superv

Rethinking Amortized Neural Representations for High-Resolution Terrain Elevation Data

Model ReleasesDGX agent

arXiv:2606.00404v1 Announce Type: new Abstract: Implicit neural representations (INRs) model a signal as a continuous coordinate-to-value function. For terrain elevation data, this supports analytic d

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete

Model ReleasesDGX agent

arXiv:2606.01532v1 Announce Type: new Abstract: Positional encoding (PE) is widely viewed as necessary for transformers to process ordered sequences: without them, the next-token map appears permutati

Revisiting Ripple Effects in Knowledge Editing through Pressure-Aware Joint Neighborhood Optimization

Model ReleasesDGX agent

arXiv:2606.01610v1 Announce Type: new Abstract: Single-edit updates in large language models can trigger ripple effects across local knowledge neighborhoods: desirable propagation to related facts and

RoboBenchMart: Benchmarking Robots in Retail Environment

Model ReleasesDGX agent

arXiv:2511.10276v2 Announce Type: replace-cross Abstract: Most existing robotic manipulation benchmarks focus on tabletop or household scenarios. While these setups have driven impressive progress, it

S-SPPO: Semantic-Calibrated Self-Play Preference Optimization

Model ReleasesDGX agent

arXiv:2606.01561v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human preferences is often formulated via Direct Preference Optimization (DPO). However, the standard Bradley

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

Model ReleasesDGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Model ReleasesDGX agent

arXiv:2503.06520v3 Announce Type: replace Abstract: Traditional methods for reasoning segmentation rely on supervised fine-tuning with categorical labels and simple descriptions, limiting its out-of-d

Silent Failures in Physical AI: A Literature Review of Runtime Action Authorization for Autonomous Systems

SafetyDGX agent

arXiv:2606.00090v1 Announce Type: cross Abstract: Physical AI systems increasingly map multimodal observations, language instructions, and learned world representations into physically consequential a

StreamingVLM: Real-Time Understanding for Infinite Video Streams

Model ReleasesDGX agent

arXiv:2510.09608v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) could power real-time assistants and autonomous agents, but they face a critical challenge: understanding near-i

← Previous
1…345346347348349…1042
Next →