AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Research

ViT^3: Unlocking Test-Time Training in Vision

DGX agent

arXiv:2512.01643v2 Announce Type: replace Abstract: Test-Time Training (TTT) has recently emerged as a promising direction for efficient sequence modeling. TTT reformulates attention operation as an o

researcharxiv-cs-cv
21 Apr 2026
Research

When More Words Say Less: Decoupling Length and Specificity in Image Description Evaluation

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.04609v2 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly used to make visual content accessible via text-based descriptions. In current systems, however, desc

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Art3D: Training-Free 3D Generation from Flat-Colored Illustration

DGX agent

arXiv:2504.10466v2 Announce Type: replace Abstract: Large-scale pre-trained image-to-3D generative models have exhibited remarkable capabilities in diverse shape generations. However, most of them str

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

AscendKernelGen: A Systematic Study of LLM-Based Kernel Generation for Neural Processing Units

DGX agent

arXiv:2601.07160v2 Announce Type: replace Abstract: To meet the ever-increasing demand for computational efficiency, Neural Processing Units (NPUs) have become critical in modern AI infrastructure. Ho

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei foo…

DGX agent

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei food. No fluff. April 28 | 5PM | Kaiyō SF Register → http://lum

model-releasesai21-labs--x
20 Apr 2026
Model Releases

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

DGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

model-releasesyann-lecun--x
20 Apr 2026
Model Releases

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

DGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

model-releaseskimi-moonshot--x
20 Apr 2026
Local Ai

kimi k2.6 is now 'available' on ollama cloud

DGX agent

Kimi K2.6 is an open-source model featuring advanced coding, long-horizon execution, and agent swarm capabilities that is now available via Ollama Cloud . The model excels in coding and agentic tools

local-air-ollama
20 Apr 2026
Applications

Modern Structure-Aware Simplicial Spatiotemporal Neural Network

DGX agent

arXiv:2604.15833v1 Announce Type: new Abstract: Spatiotemporal modeling has evolved beyond simple time series analysis to become fundamental in structural time series analysis. While current research

applicationsarxiv-cs-lg
20 Apr 2026
Model Releases

PILOT: A Promptable Interleaved Layout-aware OCR Transformer

DGX agent

arXiv:2504.03621v2 Announce Type: replace Abstract: Classical OCR pipelines decompose document reading into detection, segmentation, and recognition stages, which makes them sensitive to localization

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Reading today's open-closed performance gap

DGX agent

This article analyzes the performance differences between open-source and closed-source AI models in the current landscape, examining factors that influence their relative capabilities and market posi

researchinterconnects
20 Apr 2026
Research

SLE-FNO: Single-Layer Extensions for Task-Agnostic Continual Learning in Fourier Neural Operators

DGX agent

arXiv:2603.20410v2 Announce Type: replace Abstract: Scientific machine learning is increasingly used to build surrogate models, yet most models are trained under a restrictive assumption in which futu

researcharxiv-cs-lg
20 Apr 2026
Model Releases

Was happy with Gemma 4 Cloud, but had to change due to API Errors, GLM 5.1 spends a lot more ressources

DGX agent

A user reported satisfaction with Gemma 4 Cloud but switched to GLM 5.1 due to API errors, noting that the alternative model consumes significantly more resources. The post likely discusses the perfor

model-releasesr-ollama
20 Apr 2026
Local Ai

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it whe…

DGX agent

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it when it got cloned into Jackrongs repo, and it’s going a bit vi

local-aiclem-delangue--x
18 Apr 2026
Local Ai

Deepfake Detection Generalization with Diffusion Noise

DGX agent

arXiv:2604.14570v1 Announce Type: new Abstract: Deepfake detectors face growing challenges in generalization as new image synthesis techniques emerge. In particular, deepfakes generated by diffusion m

local-aiarxiv-cs-cv
17 Apr 2026
Research

ELMoE-3D: Leveraging Intrinsic Elasticity of MoE for Hybrid-Bonding-Enabled Self-Speculative Decoding in On-Premises Serving

DGX agent

arXiv:2604.14626v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have become the dominant architecture for large-scale language models, yet on-premises serving remains fundamentally mem

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)

DGX agent

Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn

model-releasesr-stablediffusion
17 Apr 2026
Research

Grading the Unspoken: Evaluating Tacit Reasoning in Quantum Field Theory and String Theory with LLMs

DGX agent

arXiv:2604.14188v1 Announce Type: cross Abstract: Large language models have demonstrated impressive performance across many domains of mathematics and physics. One natural question is whether such mo

researcharxiv-cs-cl
17 Apr 2026
Research

Graph-Based Alternatives to LLMs for Human Simulation

DGX agent

arXiv:2511.02135v2 Announce Type: replace Abstract: Large language models (LLMs) have become a popular approach for simulating human behaviors, yet it remains unclear if LLMs are necessary for all sim

researcharxiv-cs-cl
17 Apr 2026
Model Releases

IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation

DGX agent

arXiv:2511.01014v3 Announce Type: replace Abstract: Instruction-following is a fundamental ability of Large Language Models (LLMs), requiring their generated outputs to follow multiple constraints imp

model-releasesarxiv-cs-cl
17 Apr 2026
Tutorials

Learning temporal embeddings from electronic health records of chronic kidney disease patients

DGX agent

arXiv:2601.18675v2 Announce Type: replace Abstract: We investigate whether temporal embedding models trained on longitudinal electronic health records can learn clinically meaningful representations w

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

DGX agent

arXiv:2604.15149v1 Announce Type: new Abstract: As reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for scaling reasoning capabilities in LLMs, a new failure mode

model-releasesarxiv-cs-lg
17 Apr 2026
Agents

LM Studio 0.4.12 is out now - @Alibaba_Qwen Qwen3.6 support! - Nicer PDF exports for chats - MCP servers with OAuth work on Windows - Better…

DGX agent

LM Studio version 0.4.12 introduces support for Alibaba's Qwen 3.6 model, improved PDF export functionality for chat conversations, and fixes for MCP (Model Context Protocol) servers with OAuth authen

agentslm-studio--x
17 Apr 2026
Applications

PAGE-4D: Disentangled Pose and Geometry Estimation for VGGT-4D Perception

DGX agent

arXiv:2510.17568v5 Announce Type: replace Abstract: Recent 3D feed-forward models, such as the Visual Geometry Grounded Transformer (VGGT), have shown strong capability in inferring 3D attributes of s

applicationsarxiv-cs-cv
17 Apr 2026
Research

Q-MambaIR: Accurate Quantized Mamba for Efficient Image Restoration

DGX agent

arXiv:2503.21970v3 Announce Type: replace Abstract: State-Space Models (SSMs) have attracted considerable attention in Image Restoration (IR) due to their ability to scale linearly sequence length whi

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoring

DGX agent

arXiv:2604.14616v1 Announce Type: new Abstract: Clinical value set authoring -- the task of identifying all codes in a standardized vocabulary that define a clinical concept -- is a recurring bottlene

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Segment-Level Coherence for Robust Harmful Intent Probing in LLMs

DGX agent

arXiv:2604.14865v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly exposed to adaptive jailbreaking, particularly in high-stakes Chemical, Biological, Radiological, and Nucl

researcharxiv-cs-cl
17 Apr 2026
Model Releases

The Fourth Challenge on Image Super-Resolution (imes4) at NTIRE 2026: Benchmark Results and Method Overview

DGX agent

arXiv:2604.14558v1 Announce Type: new Abstract: This paper presents the NTIRE 2026 image super-resolution (imes4) challenge, one of the associated competitions of the NTIRE 2026 Workshop at CVPR 2026.

model-releasesarxiv-cs-cv
17 Apr 2026
Research

TokenFormer: Unify the Multi-Field and Sequential Recommendation Worlds

DGX agent

arXiv:2604.13737v1 Announce Type: cross Abstract: Recommender systems have historically developed along two largely independent paradigms: feature interaction models for modeling correlations among mu

researcharxiv-cs-ai
17 Apr 2026
Local Ai

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators

DGX agent

arXiv:2604.03307v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable success, yet they remain prone to perception-related hallucinations in fine-graine

local-aiarxiv-cs-cv
17 Apr 2026
Agents

WybeCoder: Verified Imperative Code Generation

DGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

xFODE: An Explainable Fuzzy Additive ODE Framework for System Identification

DGX agent

arXiv:2604.14883v1 Announce Type: new Abstract: Recent advances in Deep Learning (DL) have strengthened data-driven System Identification (SysID), with Neural and Fuzzy Ordinary Differential Equation

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Beyond Uniform Sampling: Synergistic Active Learning and Input Denoising for Robust Neural Operators

DGX agent

arXiv:2604.13316v1 Announce Type: new Abstract: Neural operators have emerged as fast surrogate models for physics simulations, yet they remain acutely vulnerable to adversarial perturbations, a criti

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Claude Opus 4.7 is now available in Cursor. We've found it to be impressively autonomous and more creative in its reasoning. We're launching…

DGX agent

Cursor has integrated Claude Opus 4.7 into its platform, highlighting the model's impressive autonomous capabilities and enhanced creative reasoning abilities. The announcement suggests a new feature

model-releasescursor--x
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

DGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

ExpSeek: Self-Triggered Experience Seeking for Web Agents

DGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions

DGX agent

arXiv:2509.18847v3 Announce Type: replace-cross Abstract: Tool-augmented large language models (LLMs) are usually trained with supervised imitation or coarse-grained reinforcement learning that optimi

model-releasesarxiv-cs-cl
16 Apr 2026
Research

(How) Learning Rates Regulate Catastrophic Overtraining

DGX agent

arXiv:2604.13627v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is a common first stage of LLM post-training, teaching the model to follow instructions and shaping its behavior as a hel

researcharxiv-cs-cl
16 Apr 2026
Model Releases

How WPP accelerates humanoid robot training 10x with G4 VMs

DGX agent

Editor’s note: Today we hear from Perry Nightingale, SVP of Creative AI at WPP about the workflow that cuts training time for humanoid robots from days to minutes — plus access to the open-source code

model-releasesgoogle-cloud-ai
16 Apr 2026
Model Releases

IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languages

DGX agent

arXiv:2604.13686v1 Announce Type: new Abstract: While Large Language Models (LLMs) have significantly advanced Text-to-SQL performance, existing benchmarks predominantly focus on Western contexts and

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis

DGX agent

arXiv:2604.13201v1 Announce Type: new Abstract: Large language models are emerging as scientific assistants, but evaluating their ability to reason from empirical data remains challenging. Benchmarks

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Learning the Cue or Learning the Word? Analyzing Generalization in Metaphor Detection for Verbs

DGX agent

arXiv:2604.13713v1 Announce Type: new Abstract: Metaphor detection models achieve strong benchmark performance, yet it remains unclear whether this reflects transferable generalization or lexical memo

model-releasesarxiv-cs-cl
16 Apr 2026
Applications

Lite Any Stereo: Efficient Zero-Shot Stereo Matching

DGX agent

arXiv:2511.16555v3 Announce Type: replace Abstract: Recent advances in stereo matching have focused on accuracy, often at the cost of significantly increased model size. Traditionally, the community h

applicationsarxiv-cs-cv
16 Apr 2026
Model Releases

llm-anthropic 0.25

DGX agent

Release: llm-anthropic 0.25 New model: claude-opus-4.7, which supports thinking_effort: xhigh. #66 New thinking_display and thinking_adaptive boolean options. thinking_display summarized output is cur

model-releasessimon-willison
16 Apr 2026
Model Releases

Lossless Prompt Compression via Dictionary-Encoding and In-Context Learning: Enabling Cost-Effective LLM Analysis of Repetitive Data

DGX agent

arXiv:2604.13066v1 Announce Type: new Abstract: In-context learning has established itself as an important learning paradigm for Large Language Models (LLMs). In this paper, we demonstrate that LLMs c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

MolCryst-MLIPs: A Machine-Learned Interatomic Potentials Database for Molecular Crystals

DGX agent

arXiv:2604.13897v1 Announce Type: new Abstract: We present an open Molecular Crystal (MC) database of Machine-Learned Interatomic Potentials (MLIP) called MolCryst-MLIPs. The first release comprises f

model-releasesarxiv-cs-lg
16 Apr 2026
Research

OmniTrace: A Unified Framework for Generation-Time Attribution in Omni-Modal LLMs

DGX agent

arXiv:2604.13073v1 Announce Type: new Abstract: Modern multimodal large language models (MLLMs) generate fluent responses from interleaved text, image, audio, and video inputs. However, identifying wh

researcharxiv-cs-cl
16 Apr 2026
← Previous
1…424425426427428…1371
Next →