AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,477 results
17 Apr 2026

LM Studio 0.4.12 is out now - @Alibaba_Qwen Qwen3.6 support! - Nicer PDF exports for chats - MCP servers with OAuth work on Windows - Better…

AgentsDGX agent

LM Studio version 0.4.12 introduces support for Alibaba's Qwen 3.6 model, improved PDF export functionality for chat conversations, and fixes for MCP (Model Context Protocol) servers with OAuth authen

PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D Perception

ApplicationsDGX agent

arXiv:2510.17568v5 Announce Type: replace Abstract: Recent 3D feed-forward models, such as the Visual Geometry Grounded Transformer (VGGT), have shown strong capability in inferring 3D attributes of s

Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2503.21970v3 Announce Type: replace Abstract: State-Space Models (SSMs) have attracted considerable attention in Image Restoration (IR) due to their ability to scale linearly sequence length whi

Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring

Model ReleasesDGX agent

arXiv:2604.14616v1 Announce Type: new Abstract: Clinical value set authoring -- the task of identifying all codes in a standardized vocabulary that define a clinical concept -- is a recurring bottlene

Segment-Level Coherence for Robust Harmful Intent Probing in LLMs

ResearchDGX agent

arXiv:2604.14865v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly exposed to adaptive jailbreaking, particularly in high-stakes Chemical, Biological, Radiological, and Nucl

The Fourth Challenge on Image Super-Resolution (imes4) at NTIRE 2026: Benchmark Results and Method Overview

Model ReleasesDGX agent

arXiv:2604.14558v1 Announce Type: new Abstract: This paper presents the NTIRE 2026 image super-resolution (imes4) challenge, one of the associated competitions of the NTIRE 2026 Workshop at CVPR 2026.

TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds

ResearchDGX agent

arXiv:2604.13737v1 Announce Type: cross Abstract: Recommender systems have historically developed along two largely independent paradigms: feature interaction models for modeling correlations among mu

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators

Local AiDGX agent

arXiv:2604.03307v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success, yet they remain prone to perception-related hallucinations in fine-graine

WybeCoder: Verified Imperative Code Generation

AgentsDGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

xFODE: An Explainable Fuzzy Additive ODE Framework for System Identification

Model ReleasesDGX agent

arXiv:2604.14883v1 Announce Type: new Abstract: Recent advances in Deep Learning (DL) have strengthened data-driven System Identification (SysID), with Neural and Fuzzy Ordinary Differential Equation

16 Apr 2026

Beyond Uniform Sampling: Synergistic Active Learning and Input Denoising for Robust Neural Operators

Model ReleasesDGX agent

arXiv:2604.13316v1 Announce Type: new Abstract: Neural operators have emerged as fast surrogate models for physics simulations, yet they remain acutely vulnerable to adversarial perturbations, a criti

Claude Opus 4.7 is now available in Cursor. We've found it to be impressively autonomous and more creative in its reasoning. We're launching…

Model ReleasesDGX agent

Cursor has integrated Claude Opus 4.7 into its platform, highlighting the model's impressive autonomous capabilities and enhanced creative reasoning abilities. The announcement suggests a new feature

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

Model ReleasesDGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

Model ReleasesDGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

ExpSeek: Self-Triggered Experience Seeking for Web Agents

Model ReleasesDGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions

Model ReleasesDGX agent

arXiv:2509.18847v3 Announce Type: replace-cross Abstract: Tool-augmented large language models (LLMs) are usually trained with supervised imitation or coarse-grained reinforcement learning that optimi

(How) Learning Rates Regulate Catastrophic Overtraining

ResearchDGX agent

arXiv:2604.13627v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a hel

How WPP accelerates humanoid robot training 10x with G4 VMs

Model ReleasesDGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

Model ReleasesDGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

Model ReleasesDGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

Model ReleasesDGX agent

arXiv:2604.13713v1 Announce Type: new Abstract: Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memo

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

ApplicationsDGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

llm-anthropic 0.25

Model ReleasesDGX agent

Release: llm-anthropic 0.25 New model: claude-opus-4.7, which supports thinking_effort: xhigh. #66 New thinking_display and thinking_adaptive boolean options. thinking_display summarized output is cur

Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data

Model ReleasesDGX agent

arXiv:2604.13066v1 Announce Type: new Abstract: In-context learning has established itself as an important learning paradigm for Large Language Models (LLMs). In this paper, we demonstrate that LLMs c

MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals

Model ReleasesDGX agent

arXiv:2604.13897v1 Announce Type: new Abstract: We present an open Molecular Crystal (MC) database of Machine-Learned Interatomic Potentials (MLIP) called MolCryst-MLIPs. The first release comprises f

OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs

ResearchDGX agent

arXiv:2604.13073v1 Announce Type: new Abstract: Modern multimodal large language models (MLLMs) generate fluent responses from interleaved text, image, audio, and video inputs. However, identifying wh

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management

Model ReleasesDGX agent

arXiv:2604.13531v1 Announce Type: cross Abstract: Graphical User Interface (GUI) agents show strong capabilities for automating web tasks, but existing interactive benchmarks primarily target benign,

ROBOGATE: Adaptive Failure Discovery for Safe Robot Policy Deployment via Two-Stage Boundary-Focused Sampling

Model ReleasesDGX agent

arXiv:2603.22126v3 Announce Type: replace Abstract: Deploying learned robot manipulation policies in industrial settings requires rigorous pre-deployment validation, yet exhaustive testing across high

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

Model ReleasesDGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

Model ReleasesDGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks

SafetyDGX agent

arXiv:2601.10245v2 Announce Type: replace-cross Abstract: Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to comp

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

Model ReleasesDGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

15 Apr 2026

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) …

Model ReleasesDGX agent

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) consistently picking gold patches, we were sure Claude Opus

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception

SafetyDGX agent

arXiv:2604.12255v1 Announce Type: cross Abstract: Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effec

Asymptotically Stable Gait Generation and Instantaneous Walkability Determination for Planar Almost Linear Biped with Knees

ResearchDGX agent

arXiv:2604.12274v1 Announce Type: new Abstract: A class of planar bipedal robots with unique mechanical properties has been proposed, where all links are balanced around the hip joint, preventing natu

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

SafetyDGX agent

arXiv:2604.12506v1 Announce Type: new Abstract: Recent Audio Large Language Models (AudioLLMs) exhibit a striking performance inversion: while excelling at complex reasoning tasks, they consistently u

Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

Model ReleasesDGX agent

arXiv:2603.24647v4 Announce Type: replace Abstract: The autoresearch repository enables an LLM agent to optimize hyperparameters by editing training code directly. We use it as a testbed to compare cl

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

Model ReleasesDGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse mod…

ApplicationsDGX agent

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse models. This hurts current growth On the training side, you can

DeepMind launches Gemini Robotics-ER 1.6 to meet precise physical AI demands

Model ReleasesDGX agent

Google DeepMind, Alphabet Inc.’s artificial intelligence research division, Tuesday introduced a new foundation robotics AI model designed as a significant upgrade for understanding and precise spatia

Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning

ResearchDGX agent

arXiv:2604.12483v1 Announce Type: cross Abstract: In this article, we propose the optimization of the resolution of time-frequency atoms and the regularization of fitting models to obtain better repre

Evaluating Relational Reasoning in LLMs with REL

Model ReleasesDGX agent

arXiv:2604.12176v1 Announce Type: new Abstract: Relational reasoning is the ability to infer relations that jointly bind multiple entities, attributes, or variables. This ability is central to scienti

FaCT: Faithful Concept Traces for Explaining Neural Network Decisions

Model ReleasesDGX agent

arXiv:2510.25512v2 Announce Type: replace-cross Abstract: Deep networks have shown remarkable performance across a wide range of tasks, yet getting a global concept-level understanding of how they fun

gguf import with vision / mmproj, not working!

Local AiDGX agent

This r/ollama thread addresses a known compatibility issue where importing multimodal GGUF models into Ollama using a separate mmproj (multimodal projector) file fails to enable vision capabilities. O

GPU stays sometimes at 100% usage even when done replying. Is it normal?

HardwareDGX agent

This r/ollama post addresses a commonly reported behavior where Ollama's GPU usage remains at or near 100% even after a model has finished generating a response. Certain models appear to 'hang' after

I'm going all in on Hermes (@NousResearch, @Teknium1) as my entire agent and coding stack. Six profiles. One shared self-hosted memory store…

Model ReleasesDGX agent

I'm going all in on Hermes (@NousResearch, @Teknium1) as my entire agent and coding stack. Six profiles. One shared self-hosted memory store. Zero hosted-coder dependencies. The fleet: - pmax-mousa —

Is Vibe Coding the Future? An Empirical Assessment of LLM Generated Codes for Construction Safety

Model ReleasesDGX agent

arXiv:2604.12311v1 Announce Type: cross Abstract: The emergence of vibe coding, a paradigm where non-technical users instruct Large Language Models (LLMs) to generate executable codes via natural lang

Latent Planning Emerges with Scale

Model ReleasesDGX agent

arXiv:2604.12493v1 Announce Type: cross Abstract: LLMs can perform seemingly planning-intensive tasks, like writing coherent stories or functioning code, without explicitly verbalizing a plan; however

LaV-CoT: Language-Aware Visual CoT with Multi-Aspect Reward Optimization for Real-World Multilingual VQA

Model ReleasesDGX agent

arXiv:2509.10026v4 Announce Type: replace Abstract: As large vision language models (VLMs) advance, their capabilities in multilingual visual question answering (mVQA) have significantly improved. Cha

LLM-Guided Semantic Bootstrapping for Interpretable Text Classification with Tsetlin Machines

TutorialsDGX agent

arXiv:2604.12223v1 Announce Type: cross Abstract: Pretrained language models (PLMs) like BERT provide strong semantic representations but are costly and opaque, while symbolic models such as the Tsetl

Local-Splitter: A Measurement Study of Seven Tactics for Reducing Cloud LLM Token Usage on Coding-Agent Workloads

Local AiDGX agent

arXiv:2604.12301v1 Announce Type: cross Abstract: We present a systematic measurement study of seven tactics for reducing cloud LLM token usage when a small local model can act as a triage layer in fr

Mathematics Teachers Interactions with a Multi-Agent System for Personalized Problem Generation

AgentsDGX agent

arXiv:2604.12066v1 Announce Type: new Abstract: Large language models can increasingly adapt educational tasks to learners characteristics. In the present study, we examine a multi-agent teacher-in-th

MVAdapt: Zero-Shot Multi-Vehicle Adaptation for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.11854v1 Announce Type: cross Abstract: End-to-End (E2E) autonomous driving models are usually trained and evaluated with a fixed ego-vehicle, even though their driving policy is implicitly

Operationalising the Right to be Forgotten in LLMs: A Lightweight Sequential Unlearning Framework for Privacy-Aligned Deployment in Politically Sensitive Environments

Model ReleasesDGX agent

arXiv:2604.12459v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in politically sensitive environments, where memorisation of personal data or confidential conten

Perception-Aware Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2507.06448v5 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has proven to be a highly effective strategy for endowing Large Language Models (LLMs) with ro

Polynomial Expansion Rank Adaptation: Enhancing Low-Rank Fine-Tuning with High-Order Interactions

Model ReleasesDGX agent

arXiv:2604.11841v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) is a widely used strategy for efficient fine-tuning of large language models (LLMs), but its strictly linear structure fund

RankOOD -- Class Ranking-based Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2511.19996v2 Announce Type: replace Abstract: We propose RankOOD, a rank-based Out-of-Distribution (OOD) detection approach based on training a model with the Placket-Luce loss, which is now ext

Simulation as Supervision: Mechanistic Pretraining for Scientific Discovery

SafetyDGX agent

arXiv:2507.08977v4 Announce Type: replace-cross Abstract: Scientific modeling faces a tradeoff between the interpretability of mechanistic theory and the predictive power of machine learning. While ex

SpecBranch: Speculative Decoding via Hybrid Drafting and Rollback-Aware Branch Parallelism

ApplicationsDGX agent

arXiv:2506.01979v4 Announce Type: replace-cross Abstract: Recently, speculative decoding (SD) has emerged as a promising technique to accelerate LLM inference by employing a small draft model to propo

The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break

Model ReleasesDGX agent

arXiv:2604.11978v1 Announce Type: new Abstract: Large language model (LLM) agents perform strongly on short- and mid-horizon tasks, but often break down on long-horizon tasks that require extended, in

← Previous
1…321322323324325…1042
Next →