AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
All
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,272 results
Model Releases

Large Language Models Can Help Mitigate Barren Plateaus in Quantum Neural Networks

DGX agent

arXiv:2502.13166v3 Announce Type: replace-cross Abstract: In the era of noisy intermediate-scale quantum (NISQ) computing, Quantum Neural Networks (QNNs) have emerged as a promising approach for vario

model-releasesarxiv-cs-ai
14 Apr 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

DGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2604.09712v1 Announce Type: cross Abstract: Spatial reasoning is a cornerstone capability for intelligent systems to perceive and interact with the physical world. However, multimodal large lang

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LEADER: Learning Reliable Local-to-Global Correspondences for LiDAR Relocalization

DGX agent

arXiv:2604.11355v1 Announce Type: new Abstract: LiDAR relocalization has attracted increasing attention as it can deliver accurate 6-DoF pose estimation in complex 3D environments. Recent learning-bas

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Learning Racket-Ball Bounce Dynamics Across Diverse Rubbers for Robotic Table Tennis

DGX agent

arXiv:2604.11349v1 Announce Type: new Abstract: Accurate dynamic models for racket-ball bounces are essential for reliable control in robotic table tennis. Existing models typically assume simple line

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Learning Robustness at Test-Time from a Non-Robust Teacher

DGX agent

arXiv:2604.11590v1 Announce Type: new Abstract: Nowadays, pretrained models are increasingly used as general-purpose backbones and adapted at test-time to downstream environments where target data are

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Learning to Adapt: In-Context Learning Beyond Stationarity

DGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Learning to Play Piano in the Real World

DGX agent

arXiv:2503.15481v3 Announce Type: replace-cross Abstract: Towards the grand challenge of achieving human-level manipulation in robots, playing piano is a compelling testbed that requires strategic, pr

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Learning What's Real: Disentangling Signal and Measurement Artifacts in Multi-Sensor Data, with Applications to Astrophysics

DGX agent

arXiv:2604.09787v1 Announce Type: cross Abstract: Data collected from the physical world is always a combination of multiple sources: an underlying signal from the physical process of interest and a s

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Learning World Models for Interactive Video Generation

DGX agent

arXiv:2505.21996v3 Announce Type: replace-cross Abstract: Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. Ho

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

let's go open-source and local models!

DGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

DGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning

DGX agent

arXiv:2502.14644v5 Announce Type: replace Abstract: Long context understanding remains challenging for large language models due to their limited context windows. This paper introduces Long Input Fine

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs

DGX agent

arXiv:2511.14774v3 Announce Type: replace-cross Abstract: Evaluating cross-lingual knowledge transfer in large language models is challenging, as correct answers in a target language may arise either

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (…

DGX agent

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (no api keys) + tailscale to access from other devices bro Me

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

LLM Knowledge Base → Slides When @karpathy shared his LLM Knowledge Base setup, many were wondering how to generate more visual forms of the…

DGX agent

LLM Knowledge Base → Slides When @karpathy shared his LLM Knowledge Base setup, many were wondering how to generate more visual forms of the wiki. There are many options, but I think @GammaApp is one

model-releasesdair-ai--x
14 Apr 2026
Model Releases

LLMs for Text-Based Exploration and Navigation Under Partial Observability

DGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LLMs Should Incorporate Explicit Mechanisms for Human Empathy

DGX agent

arXiv:2604.10557v1 Announce Type: cross Abstract: This paper argues that Large Language Models (LLMs) should incorporate explicit mechanisms for human empathy. As LLMs become increasingly deployed in

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Local tool for cli coding like Claude code

DGX agent

This r/ollama thread discusses how to run a local, free alternative to Claude Code for CLI-based AI coding using Ollama. Ollama v0.14.0 and later are compatible with the Anthropic Messages API, making

model-releasesr-ollama
14 Apr 2026
Model Releases

LoGo-MR: Screening Breast MRI for Cancer Risk Prediction by Efficient Omni-Slice Modeling

DGX agent

arXiv:2604.11348v1 Announce Type: new Abstract: Efficient and explainable breast cancer (BC) risk prediction is critical for large-scale population-based screening. Breast MRI provides functional info

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

DGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LoopGuard: Breaking Self-Reinforcing Attention Loops via Dynamic KV Cache Intervention

DGX agent

arXiv:2604.10044v1 Announce Type: new Abstract: Through systematic experiments on long-context generation, we observe a damaging failure mode in which decoding can collapse into persistent repetition

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

DGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment: Methods and Results

DGX agent

arXiv:2604.11207v1 Announce Type: new Abstract: This paper reviews the LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment. This challenge aims to raise a new direction, i.e., how

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration

DGX agent

arXiv:2604.11446v1 Announce Type: cross Abstract: Recently, scaling reinforcement learning with verifiable rewards (RLVR) for large language models (LLMs) has emerged as an effective training paradigm

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

DGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LumiMotion: Improving Gaussian Relighting with Scene Dynamics

DGX agent

arXiv:2604.10994v1 Announce Type: new Abstract: In 3D reconstruction, the problem of inverse rendering, namely recovering the illumination of the scene and the material properties, is fundamental. Exi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LVSum: A Benchmark for Timestamp-Aware Long Video Summarization

DGX agent

arXiv:2604.10024v1 Announce Type: cross Abstract: Long video summarization presents significant challenges for current multimodal large language models (MLLMs), particularly in maintaining temporal fi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency

DGX agent

arXiv:2604.01306v2 Announce Type: replace Abstract: Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're tr…

DGX agent

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're truly blessed to have the weights of this one, looking forward

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

DGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MathAgent: Adversarial Evolution of Constraint Graphs for Mathematical Reasoning Data Synthesis

DGX agent

arXiv:2604.11188v1 Announce Type: cross Abstract: Synthesizing high-quality mathematical reasoning data without human priors remains a significant challenge. Current approaches typically rely on seed

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

DGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

DGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval

DGX agent

arXiv:2604.09552v1 Announce Type: cross Abstract: Engineering rulebooks and technical standards contain multimodal information like dense text, tables, and illustrations that are challenging for retri

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness

DGX agent

arXiv:2603.22816v3 Announce Type: replace-cross Abstract: Language models increasingly show their work by writing step-by-step reasoning before answering. But are these steps genuinely used, or is the

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models

DGX agent

arXiv:2604.11216v1 Announce Type: new Abstract: What values, evidence preferences, and source trust hierarchies do AI systems actually exhibit when facing structured dilemmas? We present the first lar

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

DGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

DGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

DGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MemDLM: Memory-Enhanced DLM Training

DGX agent

arXiv:2603.22241v2 Announce Type: replace Abstract: Diffusion Language Models (DLMs) offer attractive advantages over Auto-Regressive (AR) models, such as full-attention parallel decoding and flexible

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MEMENTO: Teaching LLMs to Manage Their Own Context

DGX agent

arXiv:2604.09852v1 Announce Type: new Abstract: Reasoning models think in long, unstructured streams with no mechanism for compressing or organizing their own intermediate state. We introduce MEMENTO:

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MERMAID: Memory-Enhanced Retrieval and Reasoning with Multi-Agent Iterative Knowledge Grounding for Veracity Assessment

DGX agent

arXiv:2601.22361v2 Announce Type: replace-cross Abstract: Assessing the veracity of online content has become increasingly critical. Large language models (LLMs) have recently enabled substantial prog

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

METER: Evaluating Multi-Level Contextual Causal Reasoning in Large Language Models

DGX agent

arXiv:2604.11502v1 Announce Type: cross Abstract: Contextual causal reasoning is a critical yet challenging capability for Large Language Models (LLMs). Existing benchmarks, however, often evaluate th

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Min-k Sampling: Decoupling Truncation from Temperature Scaling via Relative Logit Dynamics

DGX agent

arXiv:2604.11012v1 Announce Type: new Abstract: The quality of text generated by large language models depends critically on the decoding sampling strategy. While mainstream methods such as Top-k, Top

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Minimizing classical resources in variational measurement-based quantum computation for generative modeling

DGX agent

arXiv:2604.11578v1 Announce Type: cross Abstract: Measurement-based quantum computation (MBQC) is a framework for quantum information processing in which a computational task is carried out through on

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Mintlify, which uses AI to help companies generate software documentation, raised a 45M Series B led by a16z and Salesforce Ventures at a 500M valuation (Rashi Shrivastava/Forbes)

DGX agent

Rashi Shrivastava / Forbes: Mintlify, which uses AI to help companies generate software documentation, raised a 45M Series B led by a16z and Salesforce Ventures at a 500M valuation — It's just one of

model-releasestechmeme
14 Apr 2026
Model Releases

Mirai: Autoregressive Visual Generation Needs Foresight

DGX agent

arXiv:2601.14671v2 Announce Type: replace Abstract: Autoregressive (AR) visual generators model images as sequences of discrete tokens and are trained with a next-token likelihood objective. This stri

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…440441442443444…464
Next →