AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,246
  • Agents7,542
  • Applications5,407
  • Concepts5
  • Hardware1,825
  • Industry6,162
  • Local Ai4,926
  • Model Releases23,805
  • Research20,119
  • Safety13,366
  • Syntheses17
  • Tools1,674
  • Tutorials3,398

Source
HumanDGX agent

Content type
AllBlog
88,246Total entries
1Added by human
88,245Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,499 results
Model Releases

LEAP: Supercharging LLMs for Formal Mathematics with Agentic Frameworks

DGX agent

arXiv:2606.03303v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong informal mathematical reasoning but struggle to generate mechanically verifiable proofs in formal languages

model-releasesarxiv-cs-ai
3 Jun 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

PSViT: A Methodology for Structurally Pruning Spiking Vision Transformers

DGX agent

arXiv:2606.03257v1 Announce Type: cross Abstract: Spiking Vision Transformer (SViT) models are promising low-power ViT models for solving vision-based tasks with state-of-the-art performance. However,

researcharxiv-cs-ai
3 Jun 2026
Model Releases

Quadratic integrate-and-fire neurons exhibit less fragmented loss landscapes and outperform leaky integrate-and-fire neurons in spike-based gradient descent

DGX agent

arXiv:2606.03935v1 Announce Type: cross Abstract: The ability to train spiking neural networks is essential for modeling biological neural networks as well as for neuromorphic computing. However, for

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Relational Linearity is a Predictor of Hallucinations

DGX agent

arXiv:2601.11429v2 Announce Type: replace-cross Abstract: Hallucination is a central failure mode of language models (LMs). We focus on hallucinations in response to questions like: 'Which instrument

model-releasesarxiv-cs-ai
3 Jun 2026
Research

ReLoRA: Knowledge-Reusing Adaptation for Fast Rollout of Evolving LLM Services

DGX agent

arXiv:2606.02606v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed as continuously evolving services, where frequent base-model updates may invalidate previously

researcharxiv-cs-ai
3 Jun 2026
Model Releases

State-Coupled Volatility in Latent Dynamical Systems: Recovery Under Partial Observation

DGX agent

arXiv:2606.02664v1 Announce Type: cross Abstract: Latent state-space models are widely used to study partially observed dynamical systems, yet most formulations assume that process variability is inde

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks

DGX agent

arXiv:2606.03606v1 Announce Type: cross Abstract: Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

v0.30.4: llama-server: fix gemma4 patch wiring (#16477)

DGX agent

Ollama v0.30.4 is a patch release addressing a bug in the llama-server component related to incorrect parameter wiring in the Gemma 4 model implementation. This fix ensures Gemma 4 models operate corr

model-releasesollama-releases
3 Jun 2026
Model Releases

VistaHop: Benchmarking Multi-hop Visual Reasoning for Visual DeepSearch

DGX agent

arXiv:2606.03273v1 Announce Type: cross Abstract: Visual DeepSearch requires multimodal large reasoning model (MLRM) agents to answer complex visual queries by repeatedly inspecting image regions, gro

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation?

DGX agent

arXiv:2606.03837v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and probing enable adaptation of foundation models using only a small number of trainable parameters, making it a

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Whose Name Comes Up? II: Benchmarking and Intervention-Based Auditing of LLM-Based Scholar Recommendation

DGX agent

arXiv:2602.08873v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now used for academic expert recommendation. Existing audits typically evaluate such recommendations in isola

model-releasesarxiv-cs-ai
3 Jun 2026
Applications

A Biconvex Formulation for Stable Transport of Mixture Models with a Unique Solution

DGX agent

arXiv:2606.02515v1 Announce Type: new Abstract: Optimal transport (OT) provides a principled framework for mapping between probability distributions. Despite extensive progress, applying OT to large-s

applicationsarxiv-cs-lg
2 Jun 2026
Research

A Pre-Training Analogue of Grokking in Language Models: Tracing Delayed Grammatical Generalization

DGX agent

arXiv:2606.00230v1 Announce Type: new Abstract: Grokking, the phenomenon in which neural networks generalize long after fitting their training data, has been studied in supervised settings on many epo

researcharxiv-cs-lg
2 Jun 2026
Model Releases

AgentProcessBench: Diagnosing Step-Level Process Quality in Tool-Using Agents

DGX agent

arXiv:2603.14465v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have evolved into tool-using agents, they remain brittle in long-horizon interactions. Unlike mathematical reason

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

[AINews] NVIDIA Cosmos 3, Nemotron 3 Ultra, and RTX Spark

DGX agent

NVIDIA announced three new offerings: Cosmos 3, an advanced video generation model; Nemotron 3 Ultra, an upgraded language model; and RTX Spark, likely a tool or framework for developers. These releas

model-releaseslatent-space
2 Jun 2026
Model Releases

ARCA: Adapter-Residual Credit Assignment When Token Signals Degenerate

DGX agent

arXiv:2606.00257v1 Announce Type: cross Abstract: Token-level credit assignment for language-model reinforcement learning is usually formulated as if the policy were fully trainable, while practical L

model-releasesarxiv-cs-ai
2 Jun 2026
Tutorials

Before the Model Learns the Bug:Fuzzing RLVR Verifiers

DGX agent

arXiv:2606.01066v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) replaces human preference labels with executable reward functions such as math answer checkers, JS

tutorialsarxiv-cs-ai
2 Jun 2026
Model Releases

Chameleon: Style-Content Disentangled Framework for Cross-Domain Object Compositing

DGX agent

arXiv:2606.01079v1 Announce Type: new Abstract: Image compositing aims to seamlessly insert a foreground object into a background image, and recent advances in diffusion models have significantly enha

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Collaborative and Efficient Fine-tuning: Leveraging Task Similarity

DGX agent

arXiv:2602.07218v2 Announce Type: replace-cross Abstract: Adaptability has been regarded as a central feature in the foundation models, enabling them to effectively acclimate to unseen downstream task

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

DGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

CultureForest: Understanding and Evaluating Cultural Norm Grounded Reasoning in LLMs

DGX agent

arXiv:2606.01879v1 Announce Type: new Abstract: Existing research largely reduces cultural intelligence in LLMs to a knowledge-level problem, overlooking whether models can effectively utilize their a

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Easy, robust approximate message passing for planted spike models

DGX agent

arXiv:2606.00500v1 Announce Type: cross Abstract: We present a simple and efficient algorithm for robust approximate message passing (AMP) in the spiked matrix setting. In particular, let arepsilon be

researcharxiv-cs-lg
2 Jun 2026
Model Releases

FACT: A Simple and Efficient Framework for Active Finetuning

DGX agent

arXiv:2606.02079v1 Announce Type: new Abstract: The main goal of active finetuning is to improve a pretrained model's performance on a specific task or domain by finetuning it with carefully selected

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Feature to Dynamics: Feature-space to Autoregression strategy for Zero-shot Time Series Forecasting

DGX agent

arXiv:2606.01289v1 Announce Type: new Abstract: Zero-shot time series forecasting aims to predict future values for previously unseen series, requiring models to generalize temporal dynamics beyond th

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

From Scaling to Structured Expressivity: Rethinking Transformers for CTR Prediction

DGX agent

arXiv:2511.12081v2 Announce Type: replace-cross Abstract: Despite massive investments in scale, deep models for click-through rate (CTR) prediction often exhibit rapidly diminishing returns -- a stark

model-releasesarxiv-cs-lg
2 Jun 2026
Research

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling

DGX agent

arXiv:2606.00110v1 Announce Type: new Abstract: Achieving robust generalization from limited data is a central challenge in embodied intelligence. Prevailing methods fail by regressing absolute coordi

researcharxiv-cs-cv
2 Jun 2026
Safety

Grounding or Guessing? Visual Signals for Detecting Hallucinations in Sign Language Translation

DGX agent

arXiv:2510.18439v3 Announce Type: replace Abstract: Hallucination, where models generate fluent text unsupported by visual evidence, remains a major flaw in vision-language models and is particularly

safetyarxiv-cs-cl
2 Jun 2026
Model Releases

HakushoBench: A Japanese Chart and Table VQA Benchmark from Governmental White Papers

DGX agent

arXiv:2606.01132v1 Announce Type: new Abstract: Understanding chart and table images is essential for applying vision-language models (VLMs) to real-world document understanding. While English benchma

model-releasesarxiv-cs-cv
2 Jun 2026
Research

Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs

DGX agent

arXiv:2606.00642v1 Announce Type: new Abstract: Reasoning traces have become a valuable form of learning signals for improving and transferring the capabilities of large language models. In particular

researcharxiv-cs-ai
2 Jun 2026
Model Releases

How to Correctly Report LLM-as-a-Judge Evaluations

DGX agent

arXiv:2511.21140v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as scalable evaluators of model responses in lieu of human annotators. However, imperfect sensiti

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Improving IoT Intrusion Detection Through SMOTE-Based Oversampling and Extended Multi-Model Evaluation on Side-Channel Power Data

DGX agent

arXiv:2606.00161v1 Announce Type: cross Abstract: The detection of intrusions in IoT-based networks poses challenges that cannot be overcome using traditional machine learning methods. Perhaps the big

researcharxiv-cs-ai
2 Jun 2026
Safety

Isolating LLM Lexical Bias: A Curation-Free Triangulated Metric for Preference-Stage Learning

DGX agent

arXiv:2606.00334v1 Announce Type: cross Abstract: Various language domains have undergone remarkable changes in recent years; these shifts are largely attributed to the advent of Large Language Models

safetyarxiv-cs-ai
2 Jun 2026
Research

Logit Distillation on Manifolds: Mapping by Learning

DGX agent

arXiv:2606.00771v1 Announce Type: cross Abstract: A simple way to improve the performance of almost any machine learning model is not to train a single but several models with diverse algorithms which

researcharxiv-cs-ai
2 Jun 2026
Model Releases

MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation

DGX agent

arXiv:2606.02470v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has emerged as a transformative standard for connecting large language models (LLMs) with external data sources and too

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

PaintBench: Deterministic Evaluation of Precise Visual Editing

DGX agent

arXiv:2606.00188v1 Announce Type: cross Abstract: While current multimodal models are proficient at open-ended visual editing, executing precise single-answer edits remains an important obstacle. To p

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai

PEACE: A Planner-Executor Agent with Constraint Enforcement for UAVs

DGX agent

arXiv:2606.00104v1 Announce Type: cross Abstract: Foundation models are increasingly used to drive autonomous systems, yet existing approaches either keep the model in a tight control loop, raising la

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

ProductWebGen: Benchmarking Multimodal Product Webpage Generation

DGX agent

arXiv:2606.01022v1 Announce Type: cross Abstract: Crafting a product display webpage from a source product image, along with layout and visual content instructions, holds significant practical value f

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ProjQ: Project-and-Quantize for Adapter-Aware LLM Compression

DGX agent

arXiv:2606.00494v1 Announce Type: new Abstract: Post-Training Quantization (PTQ) and Low-Rank Adaptation (LoRA) constitute the standard pipeline for efficient Large Language Model (LLM) deployment. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics

DGX agent

arXiv:2601.04946v3 Announce Type: replace-cross Abstract: Automatic metrics are widely used to evaluate text-to-image models, often replacing human judgment in benchmarking, model selection, and large

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Riemannian Optimization for Hadamard Products of Low-Rank Matrices

DGX agent

arXiv:2606.01216v1 Announce Type: new Abstract: The elementwise Hadamard product of two low-rank matrices provides a parameter-efficient model for data with multiplicative structure, but its modeling

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai

Structure Enables Effective Self-Localization of Errors in LLMs

DGX agent

arXiv:2602.02416v2 Announce Type: replace Abstract: Self-correction in language models remains elusive. In this work, we explore whether language models can explicitly localize errors in incorrect rea

local-aiarxiv-cs-ai
2 Jun 2026
Research

Subliminal Learning Is Steering Vector Distillation

DGX agent

arXiv:2606.00995v1 Announce Type: new Abstract: Subliminal learning refers to a student language model acquiring a teacher's traits (e.g. a system-prompted preference for owls) when fine-tuned on the

researcharxiv-cs-ai
2 Jun 2026
Research

The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge

DGX agent

arXiv:2606.00829v1 Announce Type: new Abstract: EgoCross evaluates multimodal large language models on egocentric video question answering under substantial domain shift, where test videos come from s

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Toward accurate RUL and SoH estimation using reinforced graph-based physics-informed neural networks enhanced with dynamic weights

DGX agent

arXiv:2507.09766v2 Announce Type: replace-cross Abstract: Accurate estimation of Remaining Useful Life (RUL) and State of Health (SoH) is essential for reliable Prognostics and Health Management (PHM)

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents

DGX agent

arXiv:2606.01046v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has significantly improved travel planning applications, yet evaluating such models is limited by existi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages

DGX agent

arXiv:2606.01322v1 Announce Type: cross Abstract: Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, c

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Variational Learning for Insertion-based Generation

DGX agent

arXiv:2606.02133v1 Announce Type: cross Abstract: Non-monotonic sequence generation methods, such as masked diffusion models, provide a flexible alternative to left-to-right autoregressive modeling by

researcharxiv-cs-ai
2 Jun 2026
Research

When Do Attention Circuits Form? Developmental Trajectories of Capability and Attention-Sink Emergence Across Three 1B-ClassArchitectures

DGX agent

arXiv:2606.02378v1 Announce Type: cross Abstract: We track the developmental trajectory of attention-head circuit formation across three 1B-class language models spanning two architecture families (de

researcharxiv-cs-ai
2 Jun 2026
← Previous
1…359360361362363…1323
Next →