AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

Japanese tech giants launch joint venture targeting physical AI for robots and machines

DGX agent

Japanese technology giants SoftBank Group Corp., Sony Corp. and NEC Corp. are teaming up with Honda Motor Co., Ltd. on a new artificial intelligence joint venture that has a single goal: to build a tr

model-releasessiliconangle
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

K-Way Energy Probes for Metacognition Reduce to Softmax in Discriminative Predictive Coding Networks

DGX agent

arXiv:2604.11011v1 Announce Type: cross Abstract: We present this as a negative result with an explanatory mechanism, not as a formal upper bound. Predictive coding networks (PCNs) admit a K-way energ

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Knowing What to Stress: A Discourse-Conditioned Text-to-Speech Benchmark

DGX agent

arXiv:2604.10580v1 Announce Type: new Abstract: Spoken meaning often depends not only on what is said, but also on which word is emphasized. The same sentence can convey correction, contrast, or clari

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Knowledge Integration in Differentiable Models: A Comparative Study of Data-Driven, Soft-Constrained, and Hard-Constrained Paradigms for Identification and Control of the Single Machine Infinite Bus System

DGX agent

arXiv:2602.09667v2 Announce Type: replace Abstract: Integrating domain knowledge into neural networks is a central challenge in scientific machine learning. Three paradigms have emerged -- data-driven

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

ks-pret-5m: a 5 million word, 12 million token kashmiri pretraining dataset

DGX agent

arXiv:2604.11066v1 Announce Type: new Abstract: We present KS-PRET-5M, the largest publicly available pretraining dataset for the Kashmiri language, comprising 5,090,244 (5.09M) words, 27,692,959 (27.

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

LABBench2: An Improved Benchmark for AI Systems Performing Biology Research

DGX agent

arXiv:2604.09554v1 Announce Type: new Abstract: Optimism for accelerating scientific discovery with AI continues to grow. Current applications of AI in scientific research range from training dedicate

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

DGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Language Prompt vs. Image Enhancement: Boosting Object Detection With CLIP in Hazy Environments

DGX agent

arXiv:2604.10637v1 Announce Type: new Abstract: Object detection in hazy environments is challenging because degraded objects are nearly invisible and their semantics are weakened by environmental noi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Large Language Models Can Help Mitigate Barren Plateaus in Quantum Neural Networks

DGX agent

arXiv:2502.13166v3 Announce Type: replace-cross Abstract: In the era of noisy intermediate-scale quantum (NISQ) computing, Quantum Neural Networks (QNNs) have emerged as a promising approach for vario

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

DGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2604.09712v1 Announce Type: cross Abstract: Spatial reasoning is a cornerstone capability for intelligent systems to perceive and interact with the physical world. However, multimodal large lang

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LEADER: Learning Reliable Local-to-Global Correspondences for LiDAR Relocalization

DGX agent

arXiv:2604.11355v1 Announce Type: new Abstract: LiDAR relocalization has attracted increasing attention as it can deliver accurate 6-DoF pose estimation in complex 3D environments. Recent learning-bas

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Learning Racket-Ball Bounce Dynamics Across Diverse Rubbers for Robotic Table Tennis

DGX agent

arXiv:2604.11349v1 Announce Type: new Abstract: Accurate dynamic models for racket-ball bounces are essential for reliable control in robotic table tennis. Existing models typically assume simple line

model-releasesarxiv-cs-ro
14 Apr 2026
Model Releases

Learning Robustness at Test-Time from a Non-Robust Teacher

DGX agent

arXiv:2604.11590v1 Announce Type: new Abstract: Nowadays, pretrained models are increasingly used as general-purpose backbones and adapted at test-time to downstream environments where target data are

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Learning to Adapt: In-Context Learning Beyond Stationarity

DGX agent

arXiv:2604.10946v1 Announce Type: new Abstract: Transformer models have become foundational across a wide range of scientific and engineering domains due to their strong empirical performance. A key c

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Learning to Play Piano in the Real World

DGX agent

arXiv:2503.15481v3 Announce Type: replace-cross Abstract: Towards the grand challenge of achieving human-level manipulation in robots, playing piano is a compelling testbed that requires strategic, pr

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Learning What's Real: Disentangling Signal and Measurement Artifacts in Multi-Sensor Data, with Applications to Astrophysics

DGX agent

arXiv:2604.09787v1 Announce Type: cross Abstract: Data collected from the physical world is always a combination of multiple sources: an underlying signal from the physical process of interest and a s

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Learning World Models for Interactive Video Generation

DGX agent

arXiv:2505.21996v3 Announce Type: replace-cross Abstract: Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. Ho

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

let's go open-source and local models!

DGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

DGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LIFT: A Novel Framework for Enhancing Long-Context Understanding of LLMs via Long Input Fine-Tuning

DGX agent

arXiv:2502.14644v5 Announce Type: replace Abstract: Long context understanding remains challenging for large language models due to their limited context windows. This paper introduces Long Input Fine

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs

DGX agent

arXiv:2511.14774v3 Announce Type: replace-cross Abstract: Evaluating cross-lingual knowledge transfer in large language models is challenging, as correct answers in a target language may arise either

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (…

DGX agent

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (no api keys) + tailscale to access from other devices bro Me

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

LLM Knowledge Base → Slides When @karpathy shared his LLM Knowledge Base setup, many were wondering how to generate more visual forms of the…

DGX agent

LLM Knowledge Base → Slides When @karpathy shared his LLM Knowledge Base setup, many were wondering how to generate more visual forms of the wiki. There are many options, but I think @GammaApp is one

model-releasesdair-ai--x
14 Apr 2026
Model Releases

LLMs for Text-Based Exploration and Navigation Under Partial Observability

DGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LLMs Should Incorporate Explicit Mechanisms for Human Empathy

DGX agent

arXiv:2604.10557v1 Announce Type: cross Abstract: This paper argues that Large Language Models (LLMs) should incorporate explicit mechanisms for human empathy. As LLMs become increasingly deployed in

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Local tool for cli coding like Claude code

DGX agent

This r/ollama thread discusses how to run a local, free alternative to Claude Code for CLI-based AI coding using Ollama. Ollama v0.14.0 and later are compatible with the Anthropic Messages API, making

model-releasesr-ollama
14 Apr 2026
Model Releases

LoGo-MR: Screening Breast MRI for Cancer Risk Prediction by Efficient Omni-Slice Modeling

DGX agent

arXiv:2604.11348v1 Announce Type: new Abstract: Efficient and explainable breast cancer (BC) risk prediction is critical for large-scale population-based screening. Breast MRI provides functional info

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LookBench: A Live and Holistic Open Benchmark for Fashion Image Retrieval

DGX agent

arXiv:2601.14706v3 Announce Type: replace Abstract: In this paper, we present LookBench (We use the term 'look' to reflect retrieval that mirrors how people shop -- finding the exact item, a close sub

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LoopGuard: Breaking Self-Reinforcing Attention Loops via Dynamic KV Cache Intervention

DGX agent

arXiv:2604.10044v1 Announce Type: new Abstract: Through systematic experiments on long-context generation, we observe a damaging failure mode in which decoding can collapse into persistent repetition

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LottieGPT: Tokenizing Vector Animation for Autoregressive Generation

DGX agent

arXiv:2604.11792v1 Announce Type: new Abstract: Despite rapid progress in video generation, existing models are incapable of producing vector animation, a dominant and highly expressive form of multim

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment: Methods and Results

DGX agent

arXiv:2604.11207v1 Announce Type: new Abstract: This paper reviews the LoViF 2026 Challenge on Human-oriented Semantic Image Quality Assessment. This challenge aims to raise a new direction, i.e., how

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration

DGX agent

arXiv:2604.11446v1 Announce Type: cross Abstract: Recently, scaling reinforcement learning with verifiable rewards (RLVR) for large language models (LLMs) has emerged as an effective training paradigm

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LRD-Net: A Lightweight Real-Centered Detection Network for Cross-Domain Face Forgery Detection

DGX agent

arXiv:2604.10862v1 Announce Type: new Abstract: The rapid advancement of diffusion-based generative models has made face forgery detection a critical challenge in digital forensics. Current detection

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LumiMotion: Improving Gaussian Relighting with Scene Dynamics

DGX agent

arXiv:2604.10994v1 Announce Type: new Abstract: In 3D reconstruction, the problem of inverse rendering, namely recovering the illumination of the scene and the material properties, is fundamental. Exi

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LVSum: A Benchmark for Timestamp-Aware Long Video Summarization

DGX agent

arXiv:2604.10024v1 Announce Type: cross Abstract: Long video summarization presents significant challenges for current multimodal large language models (MLLMs), particularly in maintaining temporal fi

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

M2-Verify: A Large-Scale Multidomain Benchmark for Checking Multimodal Claim Consistency

DGX agent

arXiv:2604.01306v2 Announce Type: replace Abstract: Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're tr…

DGX agent

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're truly blessed to have the weights of this one, looking forward

model-releasesclem-delangue--x
14 Apr 2026
Model Releases

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

DGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MathAgent: Adversarial Evolution of Constraint Graphs for Mathematical Reasoning Data Synthesis

DGX agent

arXiv:2604.11188v1 Announce Type: cross Abstract: Synthesizing high-quality mathematical reasoning data without human priors remains a significant challenge. Current approaches typically rely on seed

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

DGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

DGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval

DGX agent

arXiv:2604.09552v1 Announce Type: cross Abstract: Engineering rulebooks and technical standards contain multimodal information like dense text, tables, and illustrations that are challenging for retri

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring and curing reasoning rigidity: from decorative chain-of-thought to genuine faithfulness

DGX agent

arXiv:2603.22816v3 Announce Type: replace-cross Abstract: Language models increasingly show their work by writing step-by-step reasoning before answering. But are these steps genuinely used, or is the

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring the Authority Stack of AI Systems: Empirical Analysis of 366,120 Forced-Choice Responses Across 8 AI Models

DGX agent

arXiv:2604.11216v1 Announce Type: new Abstract: What values, evidence preferences, and source trust hierarchies do AI systems actually exhibit when facing structured dilemmas? We present the first lar

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

DGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MEDSYN: Benchmarking Multi-EviDence SYNthesis in Complex Clinical Cases for Multimodal Large Language Models

DGX agent

arXiv:2602.21950v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have shown great potential in medical applications, yet existing benchmarks inadequately capture real-world

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MedVeriSeg: Teaching MLLM-Based Medical Segmentation Models to Verify Query Validity Without Extra Training

DGX agent

arXiv:2604.10242v1 Announce Type: new Abstract: Despite recent advances in MLLM-based medical image segmentation, existing LISA-like methods cannot reliably reject false queries and often produce hall

model-releasesarxiv-cs-cv
14 Apr 2026
← Previous
1…441442443444445…465
Next →