AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Research

veScale-FSDP: Flexible and High-Performance FSDP at Scale

DGX agent

arXiv:2602.22437v3 Announce Type: replace-cross Abstract: Fully Sharded Data Parallel (FSDP), also known as Zero Redundancy Optimizer (ZeRO), is widely used for large-scale model training, because of

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Video-ToC: Video Tree-of-Cue Reasoning

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2604.20473v1 Announce Type: new Abstract: Existing Video Large Language Models (Video LLMs) struggle with complex video understanding, exhibiting limited reasoning capabilities and potential hal

model-releasesarxiv-cs-cv
23 Apr 2026
Research

Where are they looking in the operating room?

DGX agent

arXiv:2604.20574v1 Announce Type: new Abstract: Purpose: Gaze-following, the task of inferring where individuals are looking, has been widely studied in computer vision, advancing research in visual a

researcharxiv-cs-cv
23 Apr 2026
Model Releases

[AINews] OpenAI launches GPT-Image-2

DGX agent

OpenAI has launched GPT-Image-2, an advancement in their image generation capabilities. The model likely represents improvements over previous versions in areas such as image quality, prompt understan

model-releaseslatent-space
22 Apr 2026
Model Releases

Analytical Extraction of Conditional Sobol' Indices via Basis Decomposition of Polynomial Chaos Expansions

DGX agent

arXiv:2604.19165v1 Announce Type: cross Abstract: In uncertainty quantification, evaluating sensitivity measures under specific conditions (i.e., conditional Sobol' indices) is essential for systems w

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

DGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

safetyarxiv-cs-ro
22 Apr 2026
Research

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps

DGX agent

arXiv:2604.19532v1 Announce Type: cross Abstract: Tokenizing music to fit the general framework of language models is a compelling challenge, especially considering the diverse symbolic structures in

researcharxiv-cs-ai
22 Apr 2026
Model Releases

CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation

DGX agent

arXiv:2602.20409v2 Announce Type: replace Abstract: Recent vision-language models (VLMs) such as CLIP demonstrate impressive cross-modal reasoning, extending beyond images to 3D perception. Yet, these

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

DGX agent

arXiv:2604.19262v1 Announce Type: cross Abstract: Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities.

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

DGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Decoupled DiLoCo: A new frontier for resilient, distributed AI training

DGX agent

Decoupled DiLoCo is a distributed architecture that enables training of large language models across distant data centers using lower bandwidth and improved hardware resilience by dividing training in

model-releasesgoogle-deepmind
22 Apr 2026
Safety

Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation

DGX agent

arXiv:2604.19141v1 Announce Type: new Abstract: Diffusion- and flow-based models usually allocate compute uniformly across space, updating all patches with the same timestep and number of function eva

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

DGX agent

arXiv:2604.19565v1 Announce Type: cross Abstract: Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Detoxification for LLM: From Dataset Itself

DGX agent

arXiv:2604.19124v1 Announce Type: new Abstract: Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity

local-aiarxiv-cs-cl
22 Apr 2026
Research

Discrete Tilt Matching

DGX agent

arXiv:2604.18739v1 Announce Type: new Abstract: Masked diffusion large language models (dLLMs) are a promising alternative to autoregressive generation. While reinforcement learning (RL) methods have

researcharxiv-cs-lg
22 Apr 2026
Model Releases

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

DGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Energy-Weighted Flow Matching: Unlocking Continuous Normalizing Flows for Efficient and Scalable Boltzmann Sampling

DGX agent

arXiv:2509.03726v2 Announce Type: replace-cross Abstract: Sampling from unnormalized target distributions, e.g. Boltzmann distributions mu_{ext{target}}(x) propto exp(-E(x)/T), is fundamental to many

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

FB-NLL: A Feature-Based Approach to Tackle Noisy Labels in Personalized Federated Learning

DGX agent

arXiv:2604.19729v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) aims to learn multiple task-specific models rather than a single global model across heterogeneous data distributi

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion

DGX agent

arXiv:2604.19015v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models (LLMs) is obstructed by a trilemma of challenges: protecting LLMs intellectual property (IP), ensuring

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

GenerativeMPC: VLM-RAG-guided Whole-Body MPC with Virtual Impedance for Bimanual Mobile Manipulation

DGX agent

arXiv:2604.19522v1 Announce Type: new Abstract: Bimanual mobile manipulation requires a seamless integration between high-level semantic reasoning and safe, compliant physical interaction - a challeng

model-releasesarxiv-cs-ro
22 Apr 2026
Tutorials

Gradient-Based Program Synthesis with Neurally Interpreted Languages

DGX agent

arXiv:2604.18907v1 Announce Type: cross Abstract: A central challenge in program induction has long been the trade-off between symbolic and neural approaches. Symbolic methods offer compositional gene

tutorialsarxiv-cs-ai
22 Apr 2026
Model Releases

HELM: Harness-Enhanced Long-horizon Memory for Vision-Language-Action Manipulation

DGX agent

arXiv:2604.18791v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models fail systematically on long-horizon manipulation tasks despite strong short-horizon performance. We show that this

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

DGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

DGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

model-releasesclem-delangue--x
22 Apr 2026
Research

Improvements to the post-processing of weather forecasts using machine learning and feature selection

DGX agent

arXiv:2604.19340v1 Announce Type: cross Abstract: This study aims to develop and improve machine learning-based post-processing models for precipitation, temperature, and wind speed predictions using

researcharxiv-cs-lg
22 Apr 2026
Model Releases

Introducing the Google Cloud Knowledge Catalog

DGX agent

Traditional data catalogs were built as manual inventories for technical users, focusing on table structures rather than the deep context that AI agents need. When agents lack business semantics and d

model-releasesgoogle-cloud-ai
22 Apr 2026
Safety

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

DGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'her…

DGX agent

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'hermes model' and select Kimi K2.6 to try out one of the most i

model-releaseskimi-moonshot--x
22 Apr 2026
Model Releases

LePREC: Reasoning as Classification over Structured Factors for Assessing Relevance of Legal Issues

DGX agent

arXiv:2604.19464v1 Announce Type: cross Abstract: More than half of the global population struggles to meet their civil justice needs due to limited legal resources. While Large Language Models (LLMs)

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning

DGX agent

arXiv:2604.18897v1 Announce Type: new Abstract: We present a systematic empirical study of prompt engineering for formal mathematical reasoning in the context of the SAIR Equational Theories Stage 1 c

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part…

DGX agent

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part is it doesn't use VLMs or any ML models at all. It's entire

model-releasesjerry-liu--x
22 Apr 2026
Model Releases

LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

DGX agent

arXiv:2604.19445v1 Announce Type: new Abstract: This paper presents a review for the LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aimed to advance research on real-world a

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

MSDS: Deep Structural Similarity with Multiscale Representation

DGX agent

arXiv:2604.19159v1 Announce Type: new Abstract: Deep-feature-based perceptual similarity models have demonstrated strong alignment with human visual perception in Image Quality Assessment (IQA). Howev

model-releasesarxiv-cs-cv
22 Apr 2026
Tutorials

Multimodal Transformer for Sample-Aware Prediction of Metal-Organic Framework Properties

DGX agent

arXiv:2604.19383v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are a major target of machine-learning-based property prediction, yet most models assume that a single framework repre

tutorialsarxiv-cs-ai
22 Apr 2026
Applications

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

DGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

applicationsarxiv-cs-cl
22 Apr 2026
Safety

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

DGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

safetyarxiv-cs-ai
22 Apr 2026
Research

Quantum inspired qubit qutrit neural networks for real time financial forecasting

DGX agent

arXiv:2604.18838v1 Announce Type: new Abstract: This research investigates the performance and efficacy of machine learning models in stock prediction, comparing Artificial Neural Networks (ANNs), Qua

researcharxiv-cs-ai
22 Apr 2026
Local Ai

Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI Recommendation

DGX agent

arXiv:2601.04562v2 Announce Type: replace Abstract: Generative recommendation with large language models (LLMs) reframes prediction as sequence generation, yet existing LLM-based recommenders remain l

local-aiarxiv-cs-ai
22 Apr 2026
Research

Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization

DGX agent

arXiv:2604.19079v1 Announce Type: cross Abstract: Unification of automatic speech recognition (ASR) systems reduces development and maintenance costs, but training a single model to perform well in bo

researcharxiv-cs-ai
22 Apr 2026
Model Releases

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

DGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Streaming Structured Inference with Flash-SemiCRF

DGX agent

arXiv:2604.18780v1 Announce Type: new Abstract: Semi-Markov Conditional Random Fields (semi-CRFs) assign labels to segments of a sequence rather than to individual positions, enabling exact inference

researcharxiv-cs-lg
22 Apr 2026
Model Releases

TabReX : Tabular Referenceless eXplainable Evaluation

DGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

DGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

DGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

The new Gemini Enterprise: one platform for agent development, orchestration, and governance

DGX agent

The first wave of AI changed how we find information; the next wave is changing how we get work done. Today, we’re enhancing our most powerful AI tools and bringing them together under one roof. Gemin

model-releasesgoogle-cloud-ai
22 Apr 2026
Research

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

DGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c

researcharxiv-cs-lg
22 Apr 2026
Research

Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length

DGX agent

arXiv:2603.22608v2 Announce Type: replace Abstract: Users often rely on Large Language Models (LLMs) for processing multiple documents or performing analysis over a number of instances. For example, a

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

DGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

model-releasesarxiv-cs-cv
22 Apr 2026
← Previous
1…478479480481482…1371
Next →