AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,223
  • Agents7,699
  • Applications5,506
  • Concepts5
  • Hardware1,889
  • Industry6,186
  • Local Ai5,045
  • Model Releases24,499
  • Research20,615
  • Safety13,633
  • Syntheses17
  • Tools1,677
  • Tutorials3,452

Source
HumanDGX agent

Content type
AllBlog
90,223Total entries
1Added by human
90,222Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,101 results
Model Releases

Leave it to the Specialist: Repair Sparse LLMs with Sparse Fine-Tuning via Sparsity Evolution

DGX agent

arXiv:2505.24037v3 Announce Type: replace Abstract: Sparse large language models (LLMs) offer an attractive direction toward efficient deployment, but adapting them to downstream tasks remains challen

model-releasesarxiv-cs-ai
3 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MLSkip: Data Skipping for ML Filters via Lightweight Metadata

DGX agent

arXiv:2606.03946v1 Announce Type: cross Abstract: Database vendors recently released AI functions that can be used in filter predicates. As such functions often rely on costly, black-box ML models, th

model-releasesarxiv-cs-lg
3 Jun 2026
Research

PointAction: 3D Points as Universal Action Representations for Robot Control

DGX agent

arXiv:2606.03943v1 Announce Type: new Abstract: Video-Action Models (VAMs) leverage the broad visual dynamics captured by pre-trained video diffusion models, offering a promising path toward generaliz

researcharxiv-cs-ro
3 Jun 2026
Model Releases

PubTables-v2: A new large-scale dataset for full-page and multi-page table extraction

DGX agent

arXiv:2512.10888v3 Announce Type: replace Abstract: Table extraction (TE) is a key challenge in document understanding. Traditional approaches detect tables first, then recognize their structure. Rece

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Revisiting Embodied Chain-of-Thought for Generalizable Robot Manipulation

DGX agent

arXiv:2606.03784v1 Announce Type: new Abstract: Embodied chain-of-thought (CoT) aims to bridge linguistic reasoning and robotic control, but its effective form and integration strategy remain underexp

model-releasesarxiv-cs-ro
3 Jun 2026
Research

Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta Solvers

DGX agent

arXiv:2502.08834v4 Announce Type: replace-cross Abstract: Deep generative models based on neural differential equations have become state-of-the-art for many generation tasks. These models rely on ODE

researcharxiv-cs-ai
3 Jun 2026
Model Releases

scTranslation: A Comprehensive Benchmark for Single-Cell Multi-Omics Modality Translation

DGX agent

arXiv:2606.03906v1 Announce Type: new Abstract: Simultaneous measurement of multiple omics modalities in single cells enables researchers to gain a more comprehensive understanding of cellular states

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

DGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning

DGX agent

arXiv:2606.03503v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have achieved remarkable progress thanks to Reinforcement Learning with Verifiable Rewards (RLVR) on Chain-of-Thoughts (Co

model-releasesarxiv-cs-ai
3 Jun 2026
Safety

Unified Video-Action Joint Denoising for Dexterous Action and Data Generation

DGX agent

arXiv:2606.03868v1 Announce Type: new Abstract: Recent world action models leverage video foundation models by aligning broad visual-dynamics priors with executable robot actions. We revisit this alig

safetyarxiv-cs-cv
3 Jun 2026
Model Releases

AblationBench: Evaluating Automated Planning of Ablations in Empirical AI Research

DGX agent

arXiv:2507.08038v3 Announce Type: replace-cross Abstract: Language model agents are increasingly used to automate scientific research, yet evaluating their scientific contributions remains a challenge

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Announcing Spanner Graph algorithms: Google-grade intelligence for connected data

DGX agent

At Google Cloud Next, we announced the preview of graph algorithms with Spanner Graph, bringing Google Research’s state-of-the-art graph mining capabilities natively to your database. These graph inte

model-releasesgoogle-cloud-ai
2 Jun 2026
Local Ai

ArrythML: An Autoencoder-Based TinyML Approach for On-Device Arrhythmia Detection on Resource-Constrained Embedded Systems

DGX agent

arXiv:2606.02256v1 Announce Type: new Abstract: Our work presents a method for ECG segmentation and arrhythmia detection using Tiny Machine Learning (TinyML) models for real-time, on-device inference

local-aiarxiv-cs-lg
2 Jun 2026
Research

Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations

DGX agent

arXiv:2511.20295v2 Announce Type: replace Abstract: Counterfactual explanations (CFEs) are minimal and semantically meaningful modifications of the input of a model that alter the model predictions. T

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Before and After Temperature: A Distributional View of Creative LLM Generation

DGX agent

arXiv:2606.01451v1 Announce Type: new Abstract: Reference-free evaluation of large language model (LLM) creativity relies on perplexity, entropy, and top-1 margin. We show that a much stronger signal

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Benchmark Dataset for Catalysis on 2D MXenes

DGX agent

arXiv:2606.00794v1 Announce Type: cross Abstract: Merging first-principles calculations with machine learning (ML), we aim to accelerate the exploration of catalytic behaviour in novel materials. We f

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Benchmarking LLM-as-a-Judge for Long-Form Output Evaluation

DGX agent

arXiv:2606.01629v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly used for long-form generation, reliably evaluating long-form outputs has become a critical challenge. L

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardware

DGX agent

arXiv:2606.01338v1 Announce Type: new Abstract: Biopharmaceutical manufacturing organizations operate under regulatory frameworks such as FDA guidance, EU Good Manufacturing Practice (GMP), and the EU

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beware of the Batch Size: Hyperparameter Bias in Evaluating LoRA

DGX agent

arXiv:2602.09492v2 Announce Type: replace-cross Abstract: Low-rank adaptation (LoRA) is a standard approach for fine-tuning large language models, yet its many variants report conflicting empirical ga

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Beyond Objects: Contextual Synthetic Data Generation for Fine-Grained Classification

DGX agent

arXiv:2510.24078v2 Announce Type: replace Abstract: Text-to-image (T2I) models are increasingly used for synthetic dataset generation, but generating effective synthetic training data for classificati

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Beyond Semantic Understanding: Preserving Collaborative Frequency Components in LLM-based Recommendation

DGX agent

arXiv:2508.10312v2 Announce Type: replace Abstract: Recommender systems in concert with Large Language Models (LLMs) present promising avenues for generating semantically-informed recommendations. How

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention

DGX agent

arXiv:2512.10414v2 Announce Type: replace Abstract: Recently, reinforcement learning (RL) has become a common choice in enhancing the reasoning capabilities of vision-language models (VLMs). Consideri

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Bridging the Sim-to-Real Gap in Semiconductor Visual Program Synthesis via Input Binarization

DGX agent

arXiv:2606.02434v1 Announce Type: new Abstract: Precise parametric control over circuit geometry is essential for semiconductor inspection, yet obtaining sufficient real training data remains costly.

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Child-directed speech facilitates production, not comprehension, in BabyLMs

DGX agent

arXiv:2606.01045v1 Announce Type: new Abstract: Recent studies suggest that child-directed speech is not conducive to language learning in BabyLMs. However, current evaluations focus predominantly on

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Citation Grounding: Detecting and Reducing LLM Citation Hallucinations via Legal Citation Graphs

DGX agent

arXiv:2606.00898v1 Announce Type: new Abstract: Large language models systematically hallucinate legal citations -- fabricating statute references, citing repealed provisions, and confusing jurisdicti

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ClinEnv: An Interactive Multi-Stage Long Horizon EHR Environment for Agents

DGX agent

arXiv:2606.02568v1 Announce Type: new Abstract: Clinical practice is not the selection of an answer from enumerated options: a physician gathers heterogeneous information incrementally and commits to

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

ClinTutor-R1: Advancing Scalable and Robust One-to-Many Alignment in Clinical Socratic Education

DGX agent

arXiv:2512.05671v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have achieved remarkable success in dyadic (one-on-one) instruction, they face significant challenges in One-to-M

safetyarxiv-cs-cl
2 Jun 2026
Model Releases

Consistency evaluation of benchmarks used for causal discovery

DGX agent

arXiv:2606.01789v1 Announce Type: new Abstract: In graphical causal model, causal discovery aims to construct a causal graph based on numerical data and domain knowledge in plain text. However, the ev

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Decentralized Instruction Tuning: Conflict-Aware Splitting and Weight Merging

DGX agent

arXiv:2606.01717v1 Announce Type: new Abstract: Instruction tuning aligns large language models, including multimodal ones, with diverse user intents, but scaling to heterogeneous mixtures is hindered

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Disentanglement-Based Equivariant Learning for Compositional VQA

DGX agent

arXiv:2606.02168v1 Announce Type: new Abstract: Compositional visual question answering (VQA) represents a challenging yet fundamental task that requires models to comprehend novel combinations of pre

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Efficient RAG with Intent-Aware Retrieval and Semantics-Preserving Chunking

DGX agent

arXiv:2606.01240v1 Announce Type: new Abstract: The demand for powerful instruction following and reasoning capability of large language models (LLMs) has promoted rapid development of retrieval-augme

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Escaping the BLEU Trap: A Signal-Grounded Framework with Decoupled Semantic Guidance for EEG-to-Text Decoding

DGX agent

arXiv:2603.03312v3 Announce Type: replace-cross Abstract: Decoding natural language from non-invasive EEG signals is a promising yet challenging task. However, current state-of-the-art models remain c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression

DGX agent

arXiv:2606.01608v1 Announce Type: new Abstract: Most existing extreme compression methods fail to achieve an optimal rate-distortion-perception trade-off, as they typically prioritize perceptual fidel

model-releasesarxiv-cs-cv
2 Jun 2026
Research

GeistBERT: Breathing Life into German NLP

DGX agent

arXiv:2506.11903v5 Announce Type: replace Abstract: Advances in transformer-based language models have highlighted the benefits of language-specific pre-training on high-quality corpora. In this conte

researcharxiv-cs-cl
2 Jun 2026
Applications

GraspGen-X: Cross-Embodiment 6-DOF Diffusion-based Grasping

DGX agent

arXiv:2606.00998v1 Announce Type: new Abstract: We study cross-embodiment 6-DOF robot grasping. Unlike prior works, we require the model not only to generalize to novel objects / scenes but also to no

applicationsarxiv-cs-ro
2 Jun 2026
Model Releases

HalleluBERT: Let Every Token That Has Meaning Bear Its Weight

DGX agent

arXiv:2510.21372v2 Announce Type: replace Abstract: Transformer-based models have advanced NLP, yet Hebrew still lacks a RoBERTa encoder that is trained at scale and released in both base and large va

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/r…

DGX agent

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/registration If you're looking for a model to use to tackle t

model-releasescohere--x
2 Jun 2026
Model Releases

Implicit Geographic Inference in LLM Medical Triage: Language-Driven Disparities in Emergency Recommendations

DGX agent

arXiv:2606.01204v1 Announce Type: cross Abstract: We investigate whether large language models produce different medical triage recommendations for identical symptoms based solely on the language of t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages

DGX agent

arXiv:2606.01260v1 Announce Type: cross Abstract: Despite being home to more than 1300 ethnic groups and 700 indigenous languages, bias in Large Language Models has not been fully studied in Indonesia

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

IntraShuffler: A Privacy Preserving Framework for Heterogeneous DP Federated Learning

DGX agent

arXiv:2606.02563v1 Announce Type: new Abstract: Heterogeneous Differential Privacy (HDP) in Federated Learning (FL) allows clients to select individual privacy budgets (arepsilon_i) according to insti

model-releasesarxiv-cs-lg
2 Jun 2026
Safety

Latent Reasoning in TRMs is Secretly a Policy Improvement Operator

DGX agent

arXiv:2511.16886v5 Announce Type: replace-cross Abstract: Recently, small models with latent recursion have obtained promising results on complex reasoning tasks. These results are typically explained

safetyarxiv-cs-ai
2 Jun 2026
Research

Leaf Spectral Reflectance Prediction Using Multi-Head Attention Neural Networks

DGX agent

arXiv:2606.01432v1 Announce Type: new Abstract: Accurate modeling of leaf spectral reflectance from physiological and biochemical traits is essential for advancing remote sensing applications in plant

researcharxiv-cs-lg
2 Jun 2026
Research

Linguistics-Aware Non-Distortionary LLM Watermarking

DGX agent

arXiv:2606.00613v1 Announce Type: cross Abstract: Watermarking should identify language-model output without degrading quality or limiting verification to the model provider. Multilingual deployment m

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Low-Resource Safety Failures Are Action Failures, Not Representation Failures

DGX agent

arXiv:2606.01196v1 Announce Type: cross Abstract: Safety alignment learned in high-resource languages transfers poorly to low-resource languages. Models refuse harmful prompts in English but fail to r

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Make Your VLA More Robust Without More Data By Interleaving Motion Planning

DGX agent

arXiv:2606.00985v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown remarkable progress for mobile manipulation, but their performance on long-horizon tasks remains poor. Th

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution

DGX agent

arXiv:2606.00017v1 Announce Type: new Abstract: Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Navigating the Reality Gap: On-Device Continual Adaptation of ASR for Clinical Telephony

DGX agent

arXiv:2512.16401v5 Announce Type: replace Abstract: Automatic Speech Recognition (ASR) can significantly reduce documentation burden in clinical workflows, but standard models degrade sharply in real-

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

On the Generalization in Topology Optimization via Sensitivity-Conditioned Bernoulli Flow Matching

DGX agent

arXiv:2606.02179v1 Announce Type: cross Abstract: Surrogate models for topology optimization (TO) exhibit highly variable out-of-distribution (OOD) generalization under distribution shifts such as cha

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…452453454455456…1357
Next →