AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,561 results
7 Jul 2026

I released sqlite-utils 4.0, the 124th release but the first major version bump since 3.0 back in 2020 I managed to keep things backwards-co…

Model ReleasesDGX agent

I released sqlite-utils 4.0, the 124th release but the first major version bump since 3.0 back in 2020 I managed to keep things backwards-compatible all the way up to version 3.39 before the accumulat

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some s…

Model ReleasesDGX agent

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some sort of internal workspace where the model holds concepts to

ICME 2026 Grand Challenge on Cross-Scenario Defect Detection and Fine-Grained Severity Grading for High-Precision Manufacturing


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2607.04675v1 Announce Type: new Abstract: This paper presents the IEEE International Conference on Multimedia and Expo (ICME) 2026 Grand Challenge on Cross-Scenario Defect Detection and Fine-Gra

ICR-RL: Deep Reinforcement Learning via In-Context Regression

Model ReleasesDGX agent

arXiv:2509.11259v2 Announce Type: replace-cross Abstract: Recent advancements in machine learning have largely been driven by foundation models (FMs) trained on large, diverse datasets, enabling them

IDEAL-Bench: Indoor Dataset and Evaluation suite for Analyzing 3D Layout reasoning

Model ReleasesDGX agent

arXiv:2607.03614v1 Announce Type: new Abstract: Spatial question answering is the dominant paradigm for evaluating spatial intelligence in Vision-Language Models (VLMs), but it leaves a complementary

If you care about AI, and want a nuanced view, you should watch this. I promise it will be worth your time.

Model ReleasesDGX agent

If you care about AI, and want a nuanced view, you should watch this. I promise it will be worth your time. An excellent video from the @WorldSciFest featuring @GaryMarcus , in which he discusses—amon

imo this is the most impt part of anthropic's J-space paper today. it's a two-parter: 1) ant proved that they can do 'brain surgery' interve…

Model ReleasesDGX agent

imo this is the most impt part of anthropic's J-space paper today. it's a two-parter: 1) ant proved that they can do 'brain surgery' interventions into reasoning to change topics midstream* 2) THE MOD

in game mechanics, these are called a variable rewards (also, yay)

Model ReleasesDGX agent

Variable rewards are game mechanics that provide unpredictable or randomized incentives to players, designed to encourage repeated engagement through the uncertainty of outcomes. This concept is commo

Industrial3D: A Water-Treatment TLS Point Cloud Dataset and Cross-Paradigm Benchmark for MEP Scene Understanding

Model ReleasesDGX agent

arXiv:2603.28660v2 Announce Type: replace Abstract: Automated semantic understanding of dense terrestrial laser scanning (TLS) point clouds is a prerequisite for Scan-to-BIM, digital twin maintenance,

IndustryNav: Exploring Spatial Reasoning of Embodied Agents in Dynamic Industrial Navigation

Model ReleasesDGX agent

arXiv:2511.17384v2 Announce Type: replace-cross Abstract: While Visual Large Language Models (VLLMs) show great promise as embodied agents, they continue to face substantial challenges in spatial reas

InFlux++: Real and Synthetic Data for Estimating Dynamic Camera Intrinsics

Model ReleasesDGX agent

arXiv:2607.05389v1 Announce Type: new Abstract: Camera intrinsics are vital for recovering 3D structure from 2D video. However, most 3D algorithms assume fixed intrinsics throughout a video, an assump

InfraNet: Quality-Aware RGB Guidance for Efficient Infrared Object Detection

Model ReleasesDGX agent

arXiv:2607.03795v1 Announce Type: new Abstract: Robust object detection under adverse visual conditions remains a long-standing challenge for multi-modal perception systems. Existing fusion-based meth

Input Pathways Shape Few-Shot, Not Zero-Shot, Binding in Tiny Transformers: A Fully-Enumerable Study

Model ReleasesDGX agent

arXiv:2607.04926v1 Announce Type: cross Abstract: How does the way information reaches a transformer -- as symbolic tokens, a clean per-factor 'oracle' code, or an entangled perceptual vector -- shape

Integrating Neural Encoders in Bayesian Generalized Linear Mixed Models for Multimodal Data

Model ReleasesDGX agent

arXiv:2607.04647v1 Announce Type: cross Abstract: Scalable Bayesian inference for generalized linear mixed models (GLMMs) provides uncertainty-aware analysis of correlated longitudinal data, but exist

Interpretable Human-Label-Free Deep Learning for Real-Bogus Classification with Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2607.05393v1 Announce Type: cross Abstract: Time-domain surveys generate many transient candidates, making Real-Bogus classification a critical step in automated discovery pipelines. Reliable la

IPDiff: Diffusion-driven ORSI Salient Object Detection with Information Reconstruction and Multi-Prior Guidance

Model ReleasesDGX agent

arXiv:2607.03696v1 Announce Type: new Abstract: Existing Salient Object Detection in Optical Remote Sensing Image (ORSI-SOD) methods mainly adopt the static inference strategy, which uses fixed traine

IRC-Bench: Recognizing Entities from Contextual Cues in First-Person Reminiscences

Model ReleasesDGX agent

arXiv:2605.06142v2 Announce Type: replace-cross Abstract: When people recount personal memories, they often refer to people, places, and events indirectly, relying on con-textual cues rather than expl

IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation

Model ReleasesDGX agent

arXiv:2512.10730v2 Announce Type: replace Abstract: Recent advances in motion-aware large language models have shown remarkable promise for jointly learning motion understanding and generation knowled

IRIS: An Intelligent Vision-Language System for Ocular Surface Diseases via Topic Tree and Scene-Driven VQA Generation

Model ReleasesDGX agent

arXiv:2607.04344v1 Announce Type: cross Abstract: While Large Vision-Language Models (VLMs) demonstrate remarkable generic capabilities, their clinical reasoning in specialized domains like ocular sur

Is the Geometry Doing the Work? An Operating-Point Audit of Hierarchy in Hyperbolic Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.05268v1 Announce Type: new Abstract: Whether a hyperbolic representation model uses its geometry cannot be read off its curvature parameter: what matters is the dimensionless operating poin

Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

Model ReleasesDGX agent

arXiv:2503.06643v2 Announce Type: replace-cross Abstract: In this paper, we tackle a critical challenge in model evaluation: how to keep code benchmarks useful when models might have already seen them

iVISION-2DCD: A Long-Term Change Detection Dataset for Large-Scale Outdoor Construction Monitoring

Model ReleasesDGX agent

arXiv:2607.03553v1 Announce Type: new Abstract: Automation in construction is essential for reducing costs and human errors in large-scale projects. We approach the construction progress monitoring fr

JADAI: Jointly Amortizing Adaptive Design and Bayesian Inference

Model ReleasesDGX agent

arXiv:2512.22999v2 Announce Type: replace-cross Abstract: We consider problems of parameter estimation where design variables can be actively optimized to maximize information gain. To this end, we in

JavaVulBench: A Java Vulnerability Benchmark with Realistic Splits, a Unified Multi-Backend Harness, and a Leakage-Aware Evaluation Mode

Model ReleasesDGX agent

arXiv:2607.02825v1 Announce Type: cross Abstract: We release extsc{JavaVulBench}, a benchmark dataset and evaluation harness for Java vulnerability detection. The dataset contains sim30{,}600 Java met

K9-Bench: Evaluating Multimodal LLMs on Canine-Centric Videos

Model ReleasesDGX agent

arXiv:2607.02680v1 Announce Type: cross Abstract: MLLMs have shown strong zero-shot capabilities across diverse inputs such as across images, video, audio, and text. A crucial, yet underexplored, appl

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL

Model ReleasesDGX agent

arXiv:2607.03991v1 Announce Type: cross Abstract: Repeated LLM calls are the standard way to estimate how trustworthy a Text-to-SQL result is: run the pipeline multiple times, judge each SQL execution

LACE-SVD: Loss-Aware SVD with Cumulative Error Correction for LLM Compression

Model ReleasesDGX agent

arXiv:2607.03057v1 Announce Type: cross Abstract: The rapid growth in the parameter scale of large language models (LLMs) has created a strong demand for efficient compression techniques. As a hardwar

LangLoc: 'Tell Me What You See'

Model ReleasesDGX agent

arXiv:2607.05077v1 Announce Type: new Abstract: We tackle fine-grained indoor localization from natural language: given a free-form description of one's surroundings, estimate the observer's 2D positi

Language-guided Medical Image Segmentation with Target-informed Multi-level Contrastive Alignments

Model ReleasesDGX agent

arXiv:2412.13533v4 Announce Type: replace Abstract: Medical image segmentation is a fundamental task in numerous medical engineering applications. Recently, language-guided segmentation has shown prom

Last-Meter Precision Navigation for UAVs: A Diffusion-Refined Aerial Visual Servoing Approach

Model ReleasesDGX agent

arXiv:2607.04352v1 Announce Type: new Abstract: In this work, we study the last-meter precision navigation for UAVs, e.g., autonomously reaching a target within the final 10 meters using monocular vis

Latent Clarity: Bridging World-Model Kinematics to Semantic Manifolds for Video Anomaly Anticipation

Model ReleasesDGX agent

arXiv:2607.03558v1 Announce Type: cross Abstract: Continuous video anomaly detection is dominated by reactive Multiple Instance Learning (MIL) that collapses spatiotemporal features into scalar scores

Learning 3D Affordances for Blade Insertion in Cluttered Stowing

Model ReleasesDGX agent

arXiv:2607.02549v1 Announce Type: new Abstract: Many manipulation tasks require reasoning about free-space affordances: discovering volumes where an extended rigid tool can safely navigate, complement

Learning Only What Valid Adapters Can Express: Subspace-Constrained Adaptation Against Fine-Tuning Poisoning

Model ReleasesDGX agent

arXiv:2607.05300v1 Announce Type: new Abstract: Parameter-efficient fine-tuning still leaves a broad space of behavior-changing updates reachable, so a poisoned objective can be represented and optimi

Learning to Generate Multiple Objects from Dense and Occluded Layouts

Model ReleasesDGX agent

arXiv:2607.03488v1 Announce Type: new Abstract: Text-to-image diffusion models fail to generate correct object counts in dense scenes, where overlapping instances collapse into indistinguishable struc

Learning to Suppress SPAD-based LiDAR Flare

Model ReleasesDGX agent

arXiv:2607.03247v1 Announce Type: new Abstract: Single-Photon Avalanche Diode (SPAD)-based Light Detection and Ranging (LiDAR) is emerging for autonomous vehicles due to its high sensitivity and preci

Learning When to Attend: Conditional Memory Access for Long-Context LLMs

Model ReleasesDGX agent

arXiv:2603.17484v2 Announce Type: replace Abstract: Language models struggle to generalize beyond pretraining context lengths, limiting long-horizon reasoning and retrieval. Continued pretraining on l

Legible-by-Construction: Attention and End-to-End Transformers

Model ReleasesDGX agent

arXiv:2607.04319v1 Announce Type: new Abstract: A companion paper showed that a transformer's feed-forward layer can be rebuilt from explicit fuzzy set operations - intersection, set-difference, and a

Less Tokens, Better Forecasts: Sparse Residual Routing for Efficient Weather Prediction

Model ReleasesDGX agent

arXiv:2607.02829v1 Announce Type: new Abstract: Existing ViT-based weather forecasting models apply uniform computation across all spatial tokens, even though nearby atmospheric grid points often cont

LGQ: Learnable Geometric Quantization for Image Tokenization

Model ReleasesDGX agent

arXiv:2602.16086v3 Announce Type: replace Abstract: Recent collapse-free quantizers such as FSQ achieve stable training by replacing the learnable codebook with an engineered geometry: a fixed scalar

LH-AVLN: A Benchmark for Long-Horizon Audio-Visual-Language Navigation

Model ReleasesDGX agent

arXiv:2607.03920v1 Announce Type: new Abstract: Embodied navigation is moving toward long-horizon missions, yet existing long-horizon benchmarks are largely acoustically silent, and audio-visual navig

Lightweight ML-Based Automatic Sleep Staging Framework with Constrained CNN and Mamba for Small-Sample EEG Datasets

Model ReleasesDGX agent

arXiv:2607.04934v1 Announce Type: new Abstract: Automatic sleep staging is a key technology for precise diagnosis and treatment of sleep disorders as well as long-term home sleep monitoring. Portable

LILAC: Layer-Wise Independent LoRAs and Cascaded Conditioning for Multi-Concept Customization of Diffusion Models

Model ReleasesDGX agent

arXiv:2607.04801v1 Announce Type: new Abstract: Personalizing text-to-image diffusion models to render several specific subjects in a coherent image remains challenging: the model must preserve each s

LiNO: Lifting based multiresolution neural operator

Model ReleasesDGX agent

arXiv:2607.02715v1 Announce Type: new Abstract: Recently, neural operators have shown promising outcomes for learning solution operators of differential equations directly from data. This framework le

LLM-as-a-Verifier: A General-Purpose Verification Framework

Model ReleasesDGX agent

arXiv:2607.05391v1 Announce Type: new Abstract: Scaling pre-training, post-training, and test-time compute have become the central paradigms for improving the capabilities of LLMs. In this work, we id

LLM-Based Test Oracles: Source-of-Authority Taxonomy -- A Systematic Literature Review

Model ReleasesDGX agent

arXiv:2607.05031v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to produce test oracles, the part of a test that decides whether observed behavior is correct. Yet

LLM-Driven CI-CD Workflow Intelligence for Cyber Systems Engineering

Model ReleasesDGX agent

arXiv:2607.04579v1 Announce Type: cross Abstract: CI/CD workflows have become executable operational policy: they decide what gets built, tested, released, and deployed, and they mediate how maintaine

LLMs Encode Harmfulness and Refusal Separately

Model ReleasesDGX agent

arXiv:2507.11878v5 Announce Type: replace Abstract: LLMs are trained to refuse harmful instructions, but do they truly understand harmfulness beyond just refusing? Prior work has shown that LLMs' refu

Localized LoRA-MoE: Block-wise Low-Rank Experts With Adaptive Routing

Model ReleasesDGX agent

arXiv:2607.05114v1 Announce Type: cross Abstract: Large Language Models (LLMs) and high-dimensional perception networks increasingly rely on parameter-efficient fine-tuning (PEFT) to adapt to diverse

Longitudinal-Motion-Aware Lateral Control for Autonomous Vehicles: A Robust Nonlinear Control Framework

Model ReleasesDGX agent

arXiv:2607.02924v1 Announce Type: new Abstract: As autonomous vehicles (AVs) operate in increasingly dynamic traffic conditions, lateral control must be performed while longitudinal speed and accelera

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL).…

Model ReleasesDGX agent

Loop engineering is great until something breaks. Here is how I improve the reliability of my agentic loops. I use human-in-the-loop (HITL). It's easy and extremely effective. Anyone can build this. M

LuxSQA: Ask Me in Luxembourgish with TTS-Augmented Spoken Question Answering

Model ReleasesDGX agent

arXiv:2607.02763v1 Announce Type: new Abstract: Spoken Question Answering (SQA) remains largely focused on high-resource languages and carefully recorded speech, limiting the reach of speech-LLM metho

MAI-1 has no independent benchmarks yet, but the ones they released suggest it is worse than Sonnet 4.6. Not sure it is going to be great as…

Model ReleasesDGX agent

MAI-1 has no independent benchmarks yet, but the ones they released suggest it is worse than Sonnet 4.6. Not sure it is going to be great as a Copilot at Excel and Outlook, especially as there are alr

MARLIN: De Novo Molecular Structure Elucidation from Tandem Mass Spectra without a Ground-Truth Formula

Model ReleasesDGX agent

arXiv:2607.04774v1 Announce Type: new Abstract: Untargeted tandem mass spectrometry (MS/MS) detects thousands of small molecules per biological sample, yet most go unidentified because they are absent

Mask2Real-WM: Segmentation Masks as a Sim-to-Real Bridge for Controllable Dexterous World Models

Model ReleasesDGX agent

arXiv:2607.04546v1 Announce Type: cross Abstract: Action-conditioned world models allow robots to predict the future consequences of candidate actions without additional physical interaction, supporti

MatPhaseBench: A Semantics-Guided Benchmark for Materials Phase Diagrams Understanding

Model ReleasesDGX agent

arXiv:2607.02934v1 Announce Type: cross Abstract: Materials phase diagrams are a core knowledge representation in materials science, encoding temperature,composition, phase stability, and phase transf

Measuring Harness-Induced Belief Divergence in Multi-Step LLM Agents

Model ReleasesDGX agent

arXiv:2607.04528v1 Announce Type: new Abstract: Software-agent benchmarks usually report whether an agent solves a task, but the agent reaches that outcome through a harness that controls what it sees

Mechanism-level routing failure in LLMs over Lean-verified algebraic structures

Model ReleasesDGX agent

arXiv:2607.04534v1 Announce Type: new Abstract: We present an empirical study of structural routing failure in large language models (LLMs) over a formally verified algebraic corpus. The task requires

MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents

Model ReleasesDGX agent

arXiv:2607.02879v1 Announce Type: new Abstract: Current benchmarks for evaluating large language models (LLMs) in medical calculation are largely based on simplified settings, where each patient case

Medi-Gemma: A Hybrid Clinical Decision Support System Integrating Deterministic EMR Analytics and Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2607.04907v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in high-stakes clinical settings remains limited by structural hallucinations, weak deterministic reasoning over

Memory chipmaker SK hynix seeks to raise $28B through US IPO

Model ReleasesDGX agent

SK hynix Inc., the South Korean firm that rivals Samsung Electronics Co. Ltd. and Micron Technology Inc. as one of the world’s top memory chipmakers, said today it will sell almost 17.8 million shares

← Previous
1…99100101102103…377
Next →