AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Model Releases

SurgCoT: Advancing Spatiotemporal Reasoning in Surgical Videos through a Chain-of-Thought Benchmark

DGX agent

arXiv:2604.20319v1 Announce Type: new Abstract: Fine-grained spatiotemporal reasoning on surgical videos is critical, yet the capabilities of Multi-modal Large Language Models (MLLMs) in this domain r

model-releasesarxiv-cs-cv
23 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Survival of the Cheapest: Cost-Aware Hardware Adaptation for Adversarial Robustness

DGX agent

arXiv:2409.07609v2 Announce Type: replace-cross Abstract: Deploying adversarially robust machine learning systems requires continuous trade-offs between robustness, cost, and latency. We present an au

model-releasesarxiv-cs-cv
23 Apr 2026
Research

The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm

DGX agent

arXiv:2604.20665v1 Announce Type: cross Abstract: The rapid proliferation of Vision-Language Models (VLMs) is widely celebrated as the dawn of unified multimodal knowledge discovery but its foundation

researcharxiv-cs-ai
23 Apr 2026
Research

The Optical and Infrared Are Connected

DGX agent

arXiv:2503.03816v2 Announce Type: replace-cross Abstract: Galaxies are often modelled as composites of separable components with distinct spectral signatures, implying that different wavelength ranges

researcharxiv-cs-lg
23 Apr 2026
Model Releases

Tokenised Flow Matching for Hierarchical Simulation Based Inference

DGX agent

arXiv:2604.20723v1 Announce Type: cross Abstract: The cost of simulator evaluations is a key practical bottleneck for Simulation Based Inference (SBI). In hierarchical settings with shared global para

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Transparent Screening for LLM Inference and Training Impacts

DGX agent

arXiv:2604.19757v1 Announce Type: cross Abstract: This paper presents a transparent screening framework for estimating inference and training impacts of current large language models under limited obs

researcharxiv-cs-ai
23 Apr 2026
Model Releases

UCCL-Zip: Lossless Compression Supercharged GPU Communication

DGX agent

arXiv:2604.17172v2 Announce Type: replace-cross Abstract: The rapid growth of large language models (LLMs) has made GPU communication a critical bottleneck. While prior work reduces communication volu

model-releasesarxiv-cs-ai
23 Apr 2026
Research

veScale-FSDP: Flexible and High-Performance FSDP at Scale

DGX agent

arXiv:2602.22437v3 Announce Type: replace-cross Abstract: Fully Sharded Data Parallel (FSDP), also known as Zero Redundancy Optimizer (ZeRO), is widely used for large-scale model training, because of

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Video-ToC: Video Tree-of-Cue Reasoning

DGX agent

arXiv:2604.20473v1 Announce Type: new Abstract: Existing Video Large Language Models (Video LLMs) struggle with complex video understanding, exhibiting limited reasoning capabilities and potential hal

model-releasesarxiv-cs-cv
23 Apr 2026
Research

Where are they looking in the operating room?

DGX agent

arXiv:2604.20574v1 Announce Type: new Abstract: Purpose: Gaze-following, the task of inferring where individuals are looking, has been widely studied in computer vision, advancing research in visual a

researcharxiv-cs-cv
23 Apr 2026
Model Releases

[AINews] OpenAI launches GPT-Image-2

DGX agent

OpenAI has launched GPT-Image-2, an advancement in their image generation capabilities. The model likely represents improvements over previous versions in areas such as image quality, prompt understan

model-releaseslatent-space
22 Apr 2026
Model Releases

Analytical Extraction of Conditional Sobol' Indices via Basis Decomposition of Polynomial Chaos Expansions

DGX agent

arXiv:2604.19165v1 Announce Type: cross Abstract: In uncertainty quantification, evaluating sensitivity measures under specific conditions (i.e., conditional Sobol' indices) is essential for systems w

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies

DGX agent

arXiv:2604.19509v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in understanding human-object interactions, but their application to robotic sys

safetyarxiv-cs-ro
22 Apr 2026
Research

BEAT: Tokenizing and Generating Symbolic Music by Uniform Temporal Steps

DGX agent

arXiv:2604.19532v1 Announce Type: cross Abstract: Tokenizing music to fit the general framework of language models is a compelling challenge, especially considering the diverse symbolic structures in

researcharxiv-cs-ai
22 Apr 2026
Model Releases

CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation

DGX agent

arXiv:2602.20409v2 Announce Type: replace Abstract: Recent vision-language models (VLMs) such as CLIP demonstrate impressive cross-modal reasoning, extending beyond images to 3D perception. Yet, these

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks

DGX agent

arXiv:2604.19262v1 Announce Type: cross Abstract: Large language models (LLMs) are now deployed worldwide, inspiring a surge of benchmarks that measure their multilingual and multicultural abilities.

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

DGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Decoupled DiLoCo: A new frontier for resilient, distributed AI training

DGX agent

Decoupled DiLoCo is a distributed architecture that enables training of large language models across distant data centers using lower bandwidth and improved hardware resilience by dividing training in

model-releasesgoogle-deepmind
22 Apr 2026
Safety

Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation

DGX agent

arXiv:2604.19141v1 Announce Type: new Abstract: Diffusion- and flow-based models usually allocate compute uniformly across space, updating all patches with the same timestep and number of function eva

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

DGX agent

arXiv:2604.19565v1 Announce Type: cross Abstract: Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Detoxification for LLM: From Dataset Itself

DGX agent

arXiv:2604.19124v1 Announce Type: new Abstract: Existing detoxification methods for large language models mainly focus on post-training stage or inference time, while few tackle the source of toxicity

local-aiarxiv-cs-cl
22 Apr 2026
Research

Discrete Tilt Matching

DGX agent

arXiv:2604.18739v1 Announce Type: new Abstract: Masked diffusion large language models (dLLMs) are a promising alternative to autoregressive generation. While reinforcement learning (RL) methods have

researcharxiv-cs-lg
22 Apr 2026
Model Releases

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

DGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Energy-Weighted Flow Matching: Unlocking Continuous Normalizing Flows for Efficient and Scalable Boltzmann Sampling

DGX agent

arXiv:2509.03726v2 Announce Type: replace-cross Abstract: Sampling from unnormalized target distributions, e.g. Boltzmann distributions mu_{ext{target}}(x) propto exp(-E(x)/T), is fundamental to many

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

FB-NLL: A Feature-Based Approach to Tackle Noisy Labels in Personalized Federated Learning

DGX agent

arXiv:2604.19729v1 Announce Type: new Abstract: Personalized Federated Learning (PFL) aims to learn multiple task-specific models rather than a single global model across heterogeneous data distributi

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion

DGX agent

arXiv:2604.19015v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models (LLMs) is obstructed by a trilemma of challenges: protecting LLMs intellectual property (IP), ensuring

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

GenerativeMPC: VLM-RAG-guided Whole-Body MPC with Virtual Impedance for Bimanual Mobile Manipulation

DGX agent

arXiv:2604.19522v1 Announce Type: new Abstract: Bimanual mobile manipulation requires a seamless integration between high-level semantic reasoning and safe, compliant physical interaction - a challeng

model-releasesarxiv-cs-ro
22 Apr 2026
Tutorials

Gradient-Based Program Synthesis with Neurally Interpreted Languages

DGX agent

arXiv:2604.18907v1 Announce Type: cross Abstract: A central challenge in program induction has long been the trade-off between symbolic and neural approaches. Symbolic methods offer compositional gene

tutorialsarxiv-cs-ai
22 Apr 2026
Model Releases

HELM: Harness-Enhanced Long-horizon Memory for Vision-Language-Action Manipulation

DGX agent

arXiv:2604.18791v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models fail systematically on long-horizon manipulation tasks despite strong short-horizon performance. We show that this

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing

DGX agent

arXiv:2604.19071v1 Announce Type: new Abstract: Evaluating the writing capabilities of large language models (LLMs) remains a significant challenge due to the multidimensional nature of writing skills

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

DGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

model-releasesclem-delangue--x
22 Apr 2026
Research

Improvements to the post-processing of weather forecasts using machine learning and feature selection

DGX agent

arXiv:2604.19340v1 Announce Type: cross Abstract: This study aims to develop and improve machine learning-based post-processing models for precipitation, temperature, and wind speed predictions using

researcharxiv-cs-lg
22 Apr 2026
Model Releases

Introducing the Google Cloud Knowledge Catalog

DGX agent

Traditional data catalogs were built as manual inventories for technical users, focusing on table structures rather than the deep context that AI agents need. When agents lack business semantics and d

model-releasesgoogle-cloud-ai
22 Apr 2026
Safety

Investigating Counterfactual Unfairness in LLMs towards Identities through Humor

DGX agent

arXiv:2604.18729v1 Announce Type: new Abstract: Humor holds up a mirror to social perception: what we find funny often reflects who we are and how we judge others. When language models engage with hum

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'her…

DGX agent

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'hermes model' and select Kimi K2.6 to try out one of the most i

model-releaseskimi-moonshot--x
22 Apr 2026
Model Releases

LePREC: Reasoning as Classification over Structured Factors for Assessing Relevance of Legal Issues

DGX agent

arXiv:2604.19464v1 Announce Type: cross Abstract: More than half of the global population struggles to meet their civil justice needs due to limited legal resources. While Large Language Models (LLMs)

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Less Is More: Cognitive Load and the Single-Prompt Ceiling in LLM Mathematical Reasoning

DGX agent

arXiv:2604.18897v1 Announce Type: new Abstract: We present a systematic empirical study of prompt engineering for formal mathematical reasoning in the context of the SAIR Equational Theories Stage 1 c

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part…

DGX agent

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part is it doesn't use VLMs or any ML models at all. It's entire

model-releasesjerry-liu--x
22 Apr 2026
Model Releases

LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results

DGX agent

arXiv:2604.19445v1 Announce Type: new Abstract: This paper presents a review for the LoViF Challenge on Real-World All-in-One Image Restoration. The challenge aimed to advance research on real-world a

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

MSDS: Deep Structural Similarity with Multiscale Representation

DGX agent

arXiv:2604.19159v1 Announce Type: new Abstract: Deep-feature-based perceptual similarity models have demonstrated strong alignment with human visual perception in Image Quality Assessment (IQA). Howev

model-releasesarxiv-cs-cv
22 Apr 2026
Tutorials

Multimodal Transformer for Sample-Aware Prediction of Metal-Organic Framework Properties

DGX agent

arXiv:2604.19383v1 Announce Type: cross Abstract: Metal-organic frameworks (MOFs) are a major target of machine-learning-based property prediction, yet most models assume that a single framework repre

tutorialsarxiv-cs-ai
22 Apr 2026
Applications

Position: LLM Watermarking Should Align Stakeholders' Incentives for Practical Adoption

DGX agent

arXiv:2510.18333v2 Announce Type: replace-cross Abstract: Despite progress in watermarking algorithms for large language models (LLMs), real-world deployment remains limited. We argue that this gap st

applicationsarxiv-cs-cl
22 Apr 2026
Safety

Product-of-Experts Training Reduces Dataset Artifacts in Natural Language Inference

DGX agent

arXiv:2604.19069v1 Announce Type: cross Abstract: Neural NLI models overfit dataset artifacts instead of truly reasoning. A hypothesis-only model gets 57.7% in SNLI, showing strong spurious correlatio

safetyarxiv-cs-ai
22 Apr 2026
Research

Quantum inspired qubit qutrit neural networks for real time financial forecasting

DGX agent

arXiv:2604.18838v1 Announce Type: new Abstract: This research investigates the performance and efficacy of machine learning models in stock prediction, comparing Artificial Neural Networks (ANNs), Qua

researcharxiv-cs-ai
22 Apr 2026
Local Ai

Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI Recommendation

DGX agent

arXiv:2601.04562v2 Announce Type: replace Abstract: Generative recommendation with large language models (LLMs) reframes prediction as sequence generation, yet existing LLM-based recommenders remain l

local-aiarxiv-cs-ai
22 Apr 2026
Research

Reducing the Offline-Streaming Gap for Unified ASR Transducer with Consistency Regularization

DGX agent

arXiv:2604.19079v1 Announce Type: cross Abstract: Unification of automatic speech recognition (ASR) systems reduces development and maintenance costs, but training a single model to perform well in bo

researcharxiv-cs-ai
22 Apr 2026
Model Releases

SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution

DGX agent

arXiv:2604.18982v1 Announce Type: new Abstract: Social intelligence, the ability to navigate complex interpersonal interactions, presents a fundamental challenge for language agents. Training such age

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Streaming Structured Inference with Flash-SemiCRF

DGX agent

arXiv:2604.18780v1 Announce Type: new Abstract: Semi-Markov Conditional Random Fields (semi-CRFs) assign labels to segments of a sequence rather than to individual positions, enabling exact inference

researcharxiv-cs-lg
22 Apr 2026
← Previous
1…473474475476477…1358
Next →