AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
2 Jul 2026

GPTKB v1.5: A Massive Knowledge Base for Exploring Factual LLM Knowledge

Model ReleasesDGX agent

arXiv:2507.05740v2 Announce Type: replace Abstract: Language models are powerful artifacts, yet their factual knowledge is still poorly understood, and inaccessible to ad-hoc browsing and scalable sta

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache

Model ReleasesDGX agent

arXiv:2607.01065v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) with extended context windows is increasingly constrained by the linear growth of Key-Value (KV) cache me

Homogenization of ell_2-Adversarial Training in High-Dimensions: Exact Dynamics under Stochastic Gradient Descent

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.00207v1 Announce Type: cross Abstract: We develop a framework for analyzing the learning dynamics of ell_2-adversarial training of single-index models on Gaussian mixtures in the high-dimen

Leveraging Multimodality for Real-Time Classification of Transients and Variables found by the Zwicky Transient Facility

ResearchDGX agent

arXiv:2607.00228v1 Announce Type: cross Abstract: Modern time-domain surveys such as the Zwicky Transient Facility (ZTF) generate hundreds of thousands of alerts each night, making real-time decisions

Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition

Model ReleasesDGX agent

arXiv:2607.00090v1 Announce Type: cross Abstract: Urban-scale Visual Place Recognition (VPR) aims to identify the geographic location of a query image by matching it against a geo-tagged database. Whi

LV-ROVER: Multi-Stream Tesseract Voting for Maltese Paragraph OCR

Model ReleasesDGX agent

arXiv:2607.00250v1 Announce Type: new Abstract: Maltese has decent text corpora and pretrained language models, but, like many languages outside the handful with large OCR benchmarks, only a single kn

Message Passing Enables Efficient Reasoning

ResearchDGX agent

arXiv:2607.01077v1 Announce Type: new Abstract: While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is

MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation

Model ReleasesDGX agent

arXiv:2602.21397v2 Announce Type: replace Abstract: Prompt learning has become a dominant paradigm for adapting vision-language models (VLMs) such as CLIP to downstream tasks without modifying pretrai

Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification

Model ReleasesDGX agent

arXiv:2607.00259v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) seeks to improve model robustness under distribution shifts by adapting parameters using unlabeled target data. However, in

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

Model ReleasesDGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

NeuroFilter: Activation-Based Guardrails for Privacy-Conscious LLM Agents

AgentsDGX agent

arXiv:2601.14660v2 Announce Type: replace-cross Abstract: Agentic Large Language Models (LLMs) are models able to reason, plan, and execute tools over unstructured data. These abilities are enabling t

Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors

Model ReleasesDGX agent

arXiv:2603.15129v3 Announce Type: replace Abstract: We present a novel paradigm for ultra-low-bitrate image compression (ULB-IC) that exploits the ``temporal'' evolution in generative image compressio

Not your weights, not your brain!

IndustryDGX agent

This post likely discusses issues of model ownership and intellectual property rights in AI development, particularly regarding concerns about who controls trained model weights and the underlying neu

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

Model ReleasesDGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

Model ReleasesDGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

Model ReleasesDGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

Model ReleasesDGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

Model ReleasesDGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

Shapley in Context: Explaining Financial Language with Domain Expertise

ApplicationsDGX agent

arXiv:2607.00856v1 Announce Type: cross Abstract: In recent years, large language models have achieved remarkable success and have seen growing adoption in financial applications. At the same time, ex

Soft Mixture-of-Recursions: Going Deeper with Recursive Vision Transformers

Model ReleasesDGX agent

arXiv:2607.00774v1 Announce Type: new Abstract: Recent recursive Transformer studies have primarily reused shared parameters across computation steps to construct compact, parameter-efficient models.

SONIC: Spectral Optimization of Noise for Inpainting with Consistency

ResearchDGX agent

arXiv:2511.19985v3 Announce Type: replace Abstract: We propose a novel training-free method for inpainting with off-the-shelf text-to-image models. While guidance-based methods in theory allow generic

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

Model ReleasesDGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

TiRex-2: Generalizing TiRex to Multivariate Data and Streaming

ApplicationsDGX agent

arXiv:2607.01204v1 Announce Type: new Abstract: We introduce TiRex-2, a recurrent xLSTM-based time series foundation model that generalizes the univariate TiRex to multivariate forecasting with both p

Towards Metric-Agnostic Trajectory Forecasting

Model ReleasesDGX agent

arXiv:2607.01133v1 Announce Type: new Abstract: Accurate trajectory forecasting of surrounding traffic participants is a core capability for autonomous driving, enabling vehicles to anticipate behavio

Two AI Metrics Diverged: Will it Make All the Difference?

SafetyDGX agent

arXiv:2607.00913v1 Announce Type: new Abstract: As exponential compute scaling continues, will the capabilities of frontier AI models outstrip what is accessible to developers on a small fixed budget?

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

Model ReleasesDGX agent

arXiv:2503.13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans. But are these explanations faithful,

1 Jul 2026

AxDafny: Agentic Verified Code Generation in Dafny

Model ReleasesDGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

Belief Contraction in Dynamic Epistemic Logic

ResearchDGX agent

arXiv:2606.31861v1 Announce Type: cross Abstract: Dynamic epistemic logic represents belief change via model transformations induced by epistemic events. Its standard formulation (Baltag, Moss, Soleck

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

Model ReleasesDGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

Bridging Scientific Heritage: An Arabic--Russian Parallel Corpus and LLM Benchmark for Sustainable Knowledge Transfer

Model ReleasesDGX agent

arXiv:2606.30943v1 Announce Type: new Abstract: Russian and Arabic are among the major languages of scientific communication. Language barriers impede the exchange of research results between these co

Calibrating the Evaluator: Does Probability Calibration Mitigate Preference Coupling in LLM Agent Feedback Loops?

Model ReleasesDGX agent

arXiv:2606.31371v1 Announce Type: cross Abstract: When large language model (LLM) agents adapt their behavior through evaluator feedback, systematic evaluator biases propagate into the agent's learned

Can Physician Expertise Improve Machine Learning Identification of Delirium?

Model ReleasesDGX agent

arXiv:2606.30651v1 Announce Type: cross Abstract: Delirium is common in hospitalized patients and is often missed in routine care. We present a user-centered interactive machine learning (UC-iML) fram

Claude Fable 5 access restored on AI Gateway

Model ReleasesDGX agent

Vercel has restored access to Claude Fable 5 on its AI Gateway service, allowing developers to integrate the model into their applications through Vercel's platform. This restoration enables continued

CSTrader: A Testbed for Language-Grounded Trading in a Community-Driven Virtual Asset Market

Model ReleasesDGX agent

arXiv:2606.31461v1 Announce Type: new Abstract: Niche asset markets, such as Counter-Strike 2 (CS2) weapon skins, are small, volatile, and heavily driven by community discussions and platform rules. T

Disentangling Reasoning Logic to Resolve Explicit Knowledge Conflicts

Model ReleasesDGX agent

arXiv:2508.01273v3 Announce Type: replace Abstract: Explicit knowledge conflicts, occurring when retrieved contexts contain contradictory information, pose a fundamental challenge for Large Language M

Distill Once, Adapt Life-Long: Exploring Dataset Distillation for Continual Test-Time Adaptation

Model ReleasesDGX agent

arXiv:2606.20196v2 Announce Type: replace Abstract: Continual Test-Time Adaptation (CTTA) aims to maintain model performance under evolving target domains by adapting online without labeled data. Howe

Dual Sparse Aggregation Transformer for Multispectral Object Detection

Model ReleasesDGX agent

arXiv:2606.31015v1 Announce Type: new Abstract: Transformer-based approaches have obtained excellent performance in multispectral object detection tasks due to their ability to model long-range depend

Fora: From Weight-Space to Function-Space Protection in Capability-Preserving Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.31092v1 Announce Type: new Abstract: Full fine-tuning adapts large language models to new tasks but can erode capabilities they already possess. Existing remedies protect through proxies su

ForgeDrive: Bidirectional Cross-Conditioning for Unified Visual-Action Generation in Autonomous Driving

AgentsDGX agent

arXiv:2606.31226v1 Announce Type: new Abstract: World-model-based autonomous driving endows the model with the ability to understand scene evolution. Yet this promise is undermined by the prevailing i

Gemini Omni Flash can swap objects, change environments, and make targeted edits to existing videos using simple text prompts. To try this w…

Model ReleasesDGX agent

Gemini 2 Flash, Google's multimodal AI model, enables video editing capabilities including object swapping, environment changes, and targeted edits through simple text prompts. The feature allows user

Gemini Omni Flash: Image to Video https://links.comfy.org/4y3e93O

Model ReleasesDGX agent

Gemini Omni Flash is a model capability that generates video content from still images, likely integrated into or demonstrated through ComfyUI's node-based interface. This feature represents advanceme

Gemini Omni Flash: Video Edit https://links.comfy.org/445S9aM

Model ReleasesDGX agent

Gemini Omni Flash: Video Edit is a ComfyUI workflow or tool that enables video editing capabilities using Google's Gemini Omni Flash model, likely allowing users to perform AI-assisted video editing t

Gemma 4 is now nearly 90% faster on Apple Silicon with Ollama using MLX! The speedup comes from improved multi-token prediction (MTP), now o…

Model ReleasesDGX agent

Gemma 4 is now nearly 90% faster on Apple Silicon with Ollama using MLX! The speedup comes from improved multi-token prediction (MTP), now on by default for Gemma 4, with more models to come. Ollama a

Geometry-Preserving Orthonormal Initialization for Low-Rank Adaptation in RLVR

Model ReleasesDGX agent

arXiv:2606.31813v1 Announce Type: cross Abstract: Low-rank adaptation (LoRA) and its variants enable parameter-efficient fine-tuning of large language models under the supervised fine-tuning (SFT) par

Improving LLM Reasoning with Homophily-aware Structural and Semantic Text-Attributed Graph Compression

Model ReleasesDGX agent

arXiv:2601.08187v3 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated promising capabilities in Text-Attributed Graph (TAG) understanding. Recent studies typically focus o

JacobianAvatar: Temporally Consistent Semi-rigid Avatar Reconstruction from a Monocular Video

Model ReleasesDGX agent

arXiv:2606.31115v1 Announce Type: new Abstract: Generating realistic human avatars in complex motions--such as clothing dynamics--requires modeling of global and local deformations which remains chall

Measuring & Mitigating Over-Alignment for LLMs in Multilingual Criminal Law Courts

Model ReleasesDGX agent

arXiv:2606.23375v2 Announce Type: replace-cross Abstract: While the wider applicability of LLMs in the legal field is currently debated due to their reliability and the gravity of any errors, narrow u

MIRTH: Mutual-Information Reasoning with Temporal Hubs for Vision-Language-Action Agents

Model ReleasesDGX agent

arXiv:2606.31167v1 Announce Type: cross Abstract: VLA models have emerged as a powerful paradigm for transferring semantic knowledge from web-scale data to physical robotic control. However, current s

Multistage Defer Trees for Hybrid Interpretability: If at First You Can't Succeed, Tree Again

ResearchDGX agent

arXiv:2606.30995v1 Announce Type: new Abstract: Recent work has shown that well-optimized individual decision trees can match complex black box models in some settings, primarily in noisy domains. For

MultiUAV-Plat: An LLM-Oriented Platform, Benchmark and Framework for Multi-UAV Collaborative Task Planning

Model ReleasesDGX agent

arXiv:2606.31073v1 Announce Type: new Abstract: Large language models (LLMs) provide a promising interface for high-level robotic task planning, but their use in multi-UAV collaboration remains diffic

One Video, One World: Turning Monocular Video into Physical 4D Scenes

Model ReleasesDGX agent

arXiv:2606.31388v1 Announce Type: new Abstract: We introduce extbf{OVOW}, the first training-free system that reconstructs instance-level, simulation-ready 4D mesh scenes from a single monocular video

PSHuman: Photorealistic Single-image 3D Human Reconstruction using Cross-Scale Multiview Diffusion and Explicit Remeshing

Local AiDGX agent

arXiv:2409.10141v3 Announce Type: replace Abstract: Detailed and photorealistic 3D human modeling is essential for various applications and has seen tremendous progress. However, full-body reconstruct

Read more: https://ollama.com/blog/faster-gemma-4-mlx-mtp

Model ReleasesDGX agent

This post from Ollama's official X account discusses performance improvements for the Gemma 4 model, likely covering optimizations related to MLX (Machine Learning eXperimental framework) and MTP (Mul

Reference-Free Image Quality Assessment for Virtual Try-On via Human Feedback

Model ReleasesDGX agent

arXiv:2603.13057v2 Announce Type: replace Abstract: As virtual try-on (VTON) systems become increasingly important in fashion e-commerce, there is a growing need for reliable reference-free evaluation

Rethinking the Role of Feature Engineering and Learning Strategies in Few-Shot Hidden Emotion Recognition

Model ReleasesDGX agent

arXiv:2606.31249v1 Announce Type: new Abstract: In this paper, we present the solution developed by our team, XInsight Lab, which achieved first place in Track 3 of the 4th EI-MIGA-IJCAI Challenge wit

Revising RVL-CDIP: Quantifying Errors and Test-Train Overlap

Model ReleasesDGX agent

arXiv:2606.31446v1 Announce Type: new Abstract: RVL-CDIP is a popular dataset for benchmarking document classifiers. However, the dataset contains ample amounts of label errors as well as non-trivial

Scaling Storm-Resolving Atmospheric AI Simulation to the Entire Planet

Local AiDGX agent

arXiv:2606.31248v1 Announce Type: cross Abstract: Kilometer-scale convection shapes precipitation extremes, tropical organization, and cloud feedbacks, but most global atmospheric models approximate t

Security--Fidelity Tradeoffs: The Hidden Cost of Prompt Injection Defense

Model ReleasesDGX agent

arXiv:2606.30783v1 Announce Type: cross Abstract: We identify a security-fidelity tradeoff in defending LLMs against indirect prompt injection: defenses resist injected instructions largely by suppres

UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization

Model ReleasesDGX agent

arXiv:2606.31732v1 Announce Type: new Abstract: Visual-to-Code generation, which transforms scientific plots, vector graphics, and webpages into executable scripts, demands a level of pixel-precise al

We are SO back

Model ReleasesDGX agent

We are SO back Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to

← Previous
1…390391392393394…1051
Next →