AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlog
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,672 results
Research

Graph-Based Alternatives to LLMs for Human Simulation

DGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

researcharxiv-cs-cl
17 Apr 2026
Model Releases

IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2511.01014v3 Announce Type: replace Abstract: Instruction-following is a fundamental ability of Large Language Models (LLMs), requiring their generated outputs to follow multiple constraints imp

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

Learning temporal embeddings from electronic health records of chronic kidney disease patients

DGX agent

arXiv:2601.18675v2 Announce Type: replace Abstract: We investigate whether temporal embedding models trained on longitudinal electronic health records can learn clinically meaningful representations w

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

DGX agent

arXiv:2604.15149v1 Announce Type: new Abstract: As reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for scaling reasoning capabilities in LLMs, a new failure mode

model-releasesarxiv-cs-lg
17 Apr 2026
Agents

LM Studio 0.4.12 is out now - @Alibaba_Qwen Qwen3.6 support! - Nicer PDF exports for chats - MCP servers with OAuth work on Windows - Better…

DGX agent

LM Studio version 0.4.12 introduces support for Alibaba's Qwen 3.6 model, improved PDF export functionality for chat conversations, and fixes for MCP (Model Context Protocol) servers with OAuth authen

agentslm-studio--x
17 Apr 2026
Applications

PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D Perception

DGX agent

arXiv:2510.17568v5 Announce Type: replace Abstract: Recent 3D feed-forward models, such as the Visual Geometry Grounded Transformer (VGGT), have shown strong capability in inferring 3D attributes of s

applicationsarxiv-cs-cv
17 Apr 2026
Research

Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration

DGX agent

arXiv:2503.21970v3 Announce Type: replace Abstract: State-Space Models (SSMs) have attracted considerable attention in Image Restoration (IR) due to their ability to scale linearly sequence length whi

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring

DGX agent

arXiv:2604.14616v1 Announce Type: new Abstract: Clinical value set authoring -- the task of identifying all codes in a standardized vocabulary that define a clinical concept -- is a recurring bottlene

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Segment-Level Coherence for Robust Harmful Intent Probing in LLMs

DGX agent

arXiv:2604.14865v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly exposed to adaptive jailbreaking, particularly in high-stakes Chemical, Biological, Radiological, and Nucl

researcharxiv-cs-cl
17 Apr 2026
Model Releases

The Fourth Challenge on Image Super-Resolution (imes4) at NTIRE 2026: Benchmark Results and Method Overview

DGX agent

arXiv:2604.14558v1 Announce Type: new Abstract: This paper presents the NTIRE 2026 image super-resolution (imes4) challenge, one of the associated competitions of the NTIRE 2026 Workshop at CVPR 2026.

model-releasesarxiv-cs-cv
17 Apr 2026
Research

TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds

DGX agent

arXiv:2604.13737v1 Announce Type: cross Abstract: Recommender systems have historically developed along two largely independent paradigms: feature interaction models for modeling correlations among mu

researcharxiv-cs-ai
17 Apr 2026
Local Ai

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators

DGX agent

arXiv:2604.03307v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success, yet they remain prone to perception-related hallucinations in fine-graine

local-aiarxiv-cs-cv
17 Apr 2026
Agents

WybeCoder: Verified Imperative Code Generation

DGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

xFODE: An Explainable Fuzzy Additive ODE Framework for System Identification

DGX agent

arXiv:2604.14883v1 Announce Type: new Abstract: Recent advances in Deep Learning (DL) have strengthened data-driven System Identification (SysID), with Neural and Fuzzy Ordinary Differential Equation

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Beyond Uniform Sampling: Synergistic Active Learning and Input Denoising for Robust Neural Operators

DGX agent

arXiv:2604.13316v1 Announce Type: new Abstract: Neural operators have emerged as fast surrogate models for physics simulations, yet they remain acutely vulnerable to adversarial perturbations, a criti

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available in Cursor. We've found it to be impressively autonomous and more creative in its reasoning. We're launching…

DGX agent

Cursor has integrated Claude Opus 4.7 into its platform, highlighting the model's impressive autonomous capabilities and enhanced creative reasoning abilities. The announcement suggests a new feature

model-releasescursor--x
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

DGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

ExpSeek: Self-Triggered Experience Seeking for Web Agents

DGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions

DGX agent

arXiv:2509.18847v3 Announce Type: replace-cross Abstract: Tool-augmented large language models (LLMs) are usually trained with supervised imitation or coarse-grained reinforcement learning that optimi

model-releasesarxiv-cs-cl
16 Apr 2026
Research

(How) Learning Rates Regulate Catastrophic Overtraining

DGX agent

arXiv:2604.13627v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a hel

researcharxiv-cs-cl
16 Apr 2026
Model Releases

How WPP accelerates humanoid robot training 10x with G4 VMs

DGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

model-releasesgoogle-cloud-ai
16 Apr 2026
Model Releases

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

DGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

DGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

DGX agent

arXiv:2604.13713v1 Announce Type: new Abstract: Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memo

model-releasesarxiv-cs-cl
16 Apr 2026
Applications

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

applicationsarxiv-cs-cv
16 Apr 2026
Model Releases

llm-anthropic 0.25

DGX agent

Release: llm-anthropic 0.25 New model: claude-opus-4.7, which supports thinking_effort: xhigh. #66 New thinking_display and thinking_adaptive boolean options. thinking_display summarized output is cur

model-releasessimon-willison
16 Apr 2026
Model Releases

Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data

DGX agent

arXiv:2604.13066v1 Announce Type: new Abstract: In-context learning has established itself as an important learning paradigm for Large Language Models (LLMs). In this paper, we demonstrate that LLMs c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals

DGX agent

arXiv:2604.13897v1 Announce Type: new Abstract: We present an open Molecular Crystal (MC) database of Machine-Learned Interatomic Potentials (MLIP) called MolCryst-MLIPs. The first release comprises f

model-releasesarxiv-cs-lg
16 Apr 2026
Research

OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs

DGX agent

arXiv:2604.13073v1 Announce Type: new Abstract: Modern multimodal large language models (MLLMs) generate fluent responses from interleaved text, image, audio, and video inputs. However, identifying wh

researcharxiv-cs-cl
16 Apr 2026
Model Releases

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2604.13531v1 Announce Type: cross Abstract: Graphical User Interface (GUI) agents show strong capabilities for automating web tasks, but existing interactive benchmarks primarily target benign,

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ROBOGATE: Adaptive Failure Discovery for Safe Robot Policy Deployment via Two-Stage Boundary-Focused Sampling

DGX agent

arXiv:2603.22126v3 Announce Type: replace Abstract: Deploying learned robot manipulation policies in industrial settings requires rigorous pre-deployment validation, yet exhaustive testing across high

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

Seek-and-Solve: Benchmarking MLLMs for Visual Clue-Driven Reasoning in Daily Scenarios

DGX agent

arXiv:2604.14041v1 Announce Type: new Abstract: Daily scenarios are characterized by visual richness, requiring Multimodal Large Language Models (MLLMs) to filter noise and identify decisive visual cl

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

ToolOmni: Enabling Open-World Tool Use via Agentic learning with Proactive Retrieval and Grounded Execution

DGX agent

arXiv:2604.13787v1 Announce Type: new Abstract: Large Language Models (LLMs) enhance their problem-solving capability by utilizing external tools. However, in open-world scenarios with massive and evo

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

TRIM: Hybrid Inference via Targeted Stepwise Routing in Multi-Step Reasoning Tasks

DGX agent

arXiv:2601.10245v2 Announce Type: replace-cross Abstract: Multi-step reasoning tasks like mathematical problem solving are vulnerable to cascading failures, where a single incorrect step leads to comp

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs

DGX agent

arXiv:2604.06484v2 Announce Type: replace Abstract: Cultural values are expressed not only through language but also through visual scenes and everyday social practices. Yet existing evaluations of cu

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) …

DGX agent

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) consistently picking gold patches, we were sure Claude Opus

model-releasesai21-labs--x
15 Apr 2026
Safety

ARGen: Affect-Reinforced Generative Augmentation towards Vision-based Dynamic Emotion Perception

DGX agent

arXiv:2604.12255v1 Announce Type: cross Abstract: Dynamic facial expression recognition in the wild remains challenging due to data scarcity and long-tail distributions, which hinder models from effec

safetyarxiv-cs-ai
15 Apr 2026
Research

Asymptotically Stable Gait Generation and Instantaneous Walkability Determination for Planar Almost Linear Biped with Knees

DGX agent

arXiv:2604.12274v1 Announce Type: new Abstract: A class of planar bipedal robots with unique mechanical properties has been proposed, where all links are balanced around the hip joint, preventing natu

researcharxiv-cs-ro
15 Apr 2026
Safety

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

DGX agent

arXiv:2604.12506v1 Announce Type: new Abstract: Recent Audio Large Language Models (AudioLLMs) exhibit a striking performance inversion: while excelling at complex reasoning tasks, they consistently u

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

Can LLMs Beat Classical Hyperparameter Optimization Algorithms? A Study on autoresearch

DGX agent

arXiv:2603.24647v4 Announce Type: replace Abstract: The autoresearch repository enables an LLM agent to optimize hyperparameters by editing training code directly. We use it as a testbed to compare cl

model-releasesarxiv-cs-lg
15 Apr 2026
Model Releases

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

DGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

model-releasesarxiv-cs-cl
15 Apr 2026
Applications

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse mod…

DGX agent

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse models. This hurts current growth On the training side, you can

applicationsethan-mollick--x
15 Apr 2026
Model Releases

DeepMind launches Gemini Robotics-ER 1.6 to meet precise physical AI demands

DGX agent

Google DeepMind, Alphabet Inc.’s artificial intelligence research division, Tuesday introduced a new foundation robotics AI model designed as a significant upgrade for understanding and precise spatia

model-releasessiliconangle
15 Apr 2026
Research

Elastic Net Regularization and Gabor Dictionary for Classification of Heart Sound Signals using Deep Learning

DGX agent

arXiv:2604.12483v1 Announce Type: cross Abstract: In this article, we propose the optimization of the resolution of time-frequency atoms and the regularization of fitting models to obtain better repre

researcharxiv-cs-ai
15 Apr 2026
Model Releases

Evaluating Relational Reasoning in LLMs with REL

DGX agent

arXiv:2604.12176v1 Announce Type: new Abstract: Relational reasoning is the ability to infer relations that jointly bind multiple entities, attributes, or variables. This ability is central to scienti

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

FaCT: Faithful Concept Traces for Explaining Neural Network Decisions

DGX agent

arXiv:2510.25512v2 Announce Type: replace-cross Abstract: Deep networks have shown remarkable performance across a wide range of tasks, yet getting a global concept-level understanding of how they fun

model-releasesarxiv-cs-ai
15 Apr 2026
Local Ai

gguf import with vision / mmproj, not working!

DGX agent

This r/ollama thread addresses a known compatibility issue where importing multimodal GGUF models into Ollama using a separate mmproj (multimodal projector) file fails to enable vision capabilities. O

local-air-ollama
15 Apr 2026
← Previous
1…410411412413414…1327
Next →