AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
Model Releases

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents

DGX agent

arXiv:2605.14133v1 Announce Type: new Abstract: Interactive agent benchmarks face a tension between scalable construction and realistic workflow evaluation. Hand-authored tasks are expensive to extend

model-releasesarxiv-cs-ai
15 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Communication-Efficient Federated Fine-Tuning

DGX agent

arXiv:2505.04535v3 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) h

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)

DGX agent

arXiv:2511.13397v2 Announce Type: replace-cross Abstract: The remarkable progress of Vision-Language Models (VLMs) on a variety of tasks has raised interest in their application to automated driving.

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Do-Undo Bench: Reversibility for Action Understanding in Image Generation

DGX agent

arXiv:2512.13609v2 Announce Type: replace Abstract: We introduce the Do-Undo task and benchmark to address a critical gap in vision-language models: understanding and generating plausible scene transf

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring

DGX agent

arXiv:2605.14589v1 Announce Type: new Abstract: Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and comp

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents

DGX agent

arXiv:2605.14153v1 Announce Type: cross Abstract: Exploitation is not a binary event. It is a ladder of acquiring progressive capabilities, from executing a single buggy line of code to taking full co

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution

DGX agent

arXiv:2605.15138v1 Announce Type: cross Abstract: Standard unlearning evaluations measure behavioral suppression in full precision, immediately after training, despite every deployed language model be

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

HDRFace: Rethinking Face Restoration with High-Dimensional Representation

DGX agent

arXiv:2605.14821v1 Announce Type: new Abstract: Face restoration under complex degradations still remains an ill-posed inverse problem due to severe information loss. Although diffusion models benefit

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning

DGX agent

arXiv:2605.15024v1 Announce Type: new Abstract: Remote sensing image change captioning (RSICC) aims to achieve high-level semantic understanding of genuine changes occurring between bi-temporal images

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling

DGX agent

arXiv:2605.14186v1 Announce Type: new Abstract: Large language models (LLMs) often expose useful signals of self-monitoring: before solving a problem, they can estimate whether they are likely to succ

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

NodeSynth: Socially Aligned Synthetic Data for AI Evaluation

DGX agent

arXiv:2605.14381v1 Announce Type: cross Abstract: Recent advancements in generative AI facilitate large-scale synthetic data generation for model evaluation. However, without targeted approaches, thes

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Test-Time Learning with an Evolving Library

DGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Text Knows What, Tables Know When: Clinical Timeline Reconstruction via Retrieval-Augmented Multimodal Alignment

DGX agent

arXiv:2605.15168v1 Announce Type: cross Abstract: Reconstructing precise clinical timelines is essential for modeling patient trajectories and forecasting risk in complex, heterogeneous conditions lik

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale

DGX agent

arXiv:2605.15053v1 Announce Type: cross Abstract: Continually pre-training a large language model on heterogeneous text domains, without replay or task labels, has remained an unsolved architectural p

model-releasesarxiv-cs-ai
15 May 2026
Applications

What Makes Words Hard? Sakura at BEA 2026 Shared Task on Vocabulary Difficulty Prediction

DGX agent

arXiv:2605.14257v1 Announce Type: new Abstract: We describe two types of models for vocabulary difficulty prediction: a high-accuracy black-box model, which achieved the top shared task result in the

applicationsarxiv-cs-cl
15 May 2026
Tutorials

Characteristic Root Analysis and Regularization for Linear Time Series Forecasting

DGX agent

arXiv:2509.23597v5 Announce Type: replace-cross Abstract: Time series forecasting remains a critical challenge across numerous domains, yet the effectiveness of complex models often varies unpredictab

tutorialsarxiv-cs-ai
14 May 2026
Model Releases

Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2

DGX agent

arXiv:2605.12968v1 Announce Type: cross Abstract: Do large language models internally encode ontological relations in a formally verifiable algebraic structure? We introduce Algebraic Ontology Project

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CR-Net: Scaling Parameter-Efficient Training with Cross-Layer Low-Rank Structure

DGX agent

arXiv:2509.18993v3 Announce Type: replace Abstract: Low-rank architectures have become increasingly important for efficient large language model (LLM) pre-training, providing substantial reductions in

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

DocAtlas: Multilingual Document Understanding Across 80+ Languages

DGX agent

arXiv:2605.12623v1 Announce Type: cross Abstract: Multilingual document understanding remains limited for low-resource languages due to scarce training data and model-based annotation pipelines that p

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Efficient compression of neural networks and datasets

DGX agent

arXiv:2505.17469v2 Announce Type: replace-cross Abstract: Compression and generalization are fundamentally related through Solomonoff induction and the minimum description length principle (MDL), whic

model-releasesarxiv-cs-ai
14 May 2026
Tutorials

How Do Transformers Learn to Associate Tokens: Gradient Leading Terms Bring Mechanistic Interpretability

DGX agent

arXiv:2601.19208v2 Announce Type: replace-cross Abstract: Semantic associations such as the link between 'bird' and 'flew' are foundational for language modeling as they enable models to go beyond mem

tutorialsarxiv-cs-lg
14 May 2026
Model Releases

Identifying the nonlinear string dynamics with port-Hamiltonian neural networks

DGX agent

arXiv:2605.12785v1 Announce Type: new Abstract: Hybrid machine learning combines physical knowledge with data-driven models to enhance interpretability and performance. In this context, Port-Hamiltoni

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

DGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Many-Shot CoT-ICL: Making In-Context Learning Truly Learn

DGX agent

arXiv:2605.13511v1 Announce Type: cross Abstract: In-context learning (ICL) adapts large language models (LLMs) to new tasks by conditioning on demonstrations in the prompt without parameter updates.

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Meta observation: DeepSeek is still king of the active-parameter ratio

DGX agent

DeepSeek maintains the highest efficiency in terms of active parameters relative to total model size, outperforming competitors in the ratio of parameters actually used during inference versus total t

model-releasessebastian-raschka--x
14 May 2026
Model Releases

MindVLA-U1: VLA Beats VA with Unified Streaming Architecture for Autonomous Driving

DGX agent

arXiv:2605.12624v1 Announce Type: cross Abstract: Autonomous driving has progressed from modular pipelines toward end-to-end unification, and Vision-Language-Action (VLA) models are a natural extensio

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

PanoWorld: Towards Spatial Supersensing in 360^irc Panorama World

DGX agent

arXiv:2605.13169v1 Announce Type: cross Abstract: Multimodal large laboratory models (MLLMs) still struggle with spatial understanding under the dominant perspective-image paradigm, which inherits the

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Reasoning to Edit: Hypothetical Instruction-Based Image Editing with Visual Reasoning

DGX agent

arXiv:2507.01908v3 Announce Type: replace Abstract: Instruction-based image editing (IIE) has advanced rapidly with the success of diffusion models. However, existing efforts primarily focus on simple

model-releasesarxiv-cs-cv
14 May 2026
Applications

Representing Higher-Order Networks: A Survey of Graph-Based Frameworks

DGX agent

arXiv:2605.12509v1 Announce Type: cross Abstract: Many real-world phenomena are naturally modeled by graphs and networks. However, classical graph models are often limited to pairwise interactions and

applicationsarxiv-cs-ai
14 May 2026
Local Ai

Ring-2.6-1T Open sourced today! Soooo looking forward to trying it on Ollama!

DGX agent

Ring-2.6-1T is a trillion-parameter flagship reasoning model designed for real-world complex task scenarios, now available as an open-source model. The model features about 63B activated parameters pe

local-air-ollama
14 May 2026
Model Releases

Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation

DGX agent

arXiv:2605.12953v1 Announce Type: cross Abstract: Language-guided segmentation transcends the scope limitations of traditional semantic segmentation, enabling models to segment arbitrary target region

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Tighter Learning Guarantees on Digital Computers via Concentration of Measure on Finite Spaces

DGX agent

arXiv:2402.05576v4 Announce Type: replace Abstract: Machine learning models with inputs in a Euclidean space R^d, when implemented on digital computers, generalize, and their generalization gap conver

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Vividh-ASR: A Complexity-Tiered Benchmark and Optimization Dynamics for Robust Indic Speech Recognition

DGX agent

arXiv:2605.13087v1 Announce Type: cross Abstract: Fine-tuning multilingual ASR models like Whisper for low-resource languages often improves read speech but degrades spontaneous audio performance, a p

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

A compilation of the open-source LoRAs for LTX 2.3 - released in May

DGX agent

LTX-2.3 is an open-source video generation model released in January 2026 that supports LoRA fine-tuning for customizing styles, characters, and use cases. The Reddit post compiles available LTX-2.3 m

model-releasesr-stablediffusion
13 May 2026
Model Releases

CAD-feature enhanced machine learning for manufacturing effort estimation on sheet metal bending parts

DGX agent

arXiv:2605.12266v1 Announce Type: new Abstract: Graph-based machine learning has emerged as a promising approach for manufacturability analysis by learning directly from CAD models represented as Boun

model-releasesarxiv-cs-cv
13 May 2026
Safety

Can Graphs Help Vision SSMs See Better?

DGX agent

arXiv:2605.11300v1 Announce Type: new Abstract: Vision state space models inherit the efficiency and long-range modeling ability of Mamba-style selective scans. However, their performance depends crit

safetyarxiv-cs-cv
13 May 2026
Model Releases

Crash Assessment via Mesh-Based Graph Neural Networks and Physics-Aware Attention

DGX agent

arXiv:2605.11784v1 Announce Type: cross Abstract: Full-vehicle crash simulations are computationally expensive, limiting their use in iterative design exploration. This work investigates learned hybri

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Google Named a Leader in the Gartner® Magic Quadrant™ for AI Application Development Platforms: Mid-cycle update

DGX agent

May 2026 update: We’ve refreshed this post to reflect our mid-cycle positioning and the evolution of our platform since the report was first published last November. Last fall, Google was recognized a

model-releasesgoogle-cloud-ai
13 May 2026
Model Releases

Keeping Score: Efficiency Improvements in Neural Likelihood Surrogate Training via Score-Augmented Loss Functions

DGX agent

arXiv:2605.12118v1 Announce Type: cross Abstract: For stochastic process models, parameter inference is often severely bottlenecked by computationally expensive likelihood functions. Simulation-based

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

LatentHDR: Decoupling Exposure from Diffusion via Conditional Latent-to-Latent Mapping for Text/Image-to-Panoramic HDR

DGX agent

arXiv:2605.11115v1 Announce Type: new Abstract: High Dynamic Range (HDR) generation remains challenging for generative models, which are largely limited to low dynamic range outputs. Recent diffusionb

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Learning to Foresee: Unveiling the Unlocking Efficiency of On-Policy Distillation

DGX agent

arXiv:2605.11739v1 Announce Type: new Abstract: On-policy distillation (OPD) has emerged as an efficient post-training paradigm for large language models. However, existing studies largely attribute t

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Long Story Short: Disentangling Compositionality and Long-Caption Understanding in Contrastive VLMs

DGX agent

arXiv:2509.19207v2 Announce Type: replace Abstract: Contrastive vision-language models (VLMs) have made significant progress in binding visual and textual information, yet understanding long, composit

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Measuring Five-Nines Reliability: Sample-Efficient LLM Evaluation in Saturated Benchmarks

DGX agent

arXiv:2605.11209v1 Announce Type: new Abstract: While existing benchmarks demonstrate the near-perfect performance of large language models (LLMs) on various tasks, this apparent saturation often obsc

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering

DGX agent

arXiv:2605.12361v1 Announce Type: new Abstract: Evaluating large language models (LLMs) in the biomedical domain requires benchmarks that can distinguish reasoning from pattern matching and remain dis

model-releasesarxiv-cs-cl
13 May 2026
Research

On Predicting the Post-training Potential of Pre-trained LLMs

DGX agent

arXiv:2605.11978v1 Announce Type: new Abstract: The performance of Large Language Models (LLMs) on downstream tasks is fundamentally constrained by the capabilities acquired during pre-training. Howev

researcharxiv-cs-cl
13 May 2026
Model Releases

Overtrained, Not Misaligned

DGX agent

arXiv:2605.12199v1 Announce Type: new Abstract: Emergent misalignment (EM), where fine-tuning on a narrow task (like insecure code) causes broad misalignment across unrelated domains, was first demons

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

PreScam: A Benchmark for Predicting Scam Progression from Early Conversations

DGX agent

arXiv:2605.12243v1 Announce Type: new Abstract: Conversational scams, such as romance and investment scams, are emerging as a major form of online fraud. Unlike one-shot scam lures such as fake lotter

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Scaling Laws and Tradeoffs in Recurrent Networks of Expressive Neurons

DGX agent

arXiv:2605.12049v1 Announce Type: new Abstract: Cortical neurons are complex, multi-timescale processors wired into recurrent circuits, shaped by long evolutionary pressure under stringent biological

model-releasesarxiv-cs-lg
13 May 2026
← Previous
1…401402403404405…1326
Next →