AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,489 results
Research

The Tell-Tale Norm: ell_2 Magnitude as a Signal for Reasoning Dynamics in Large Language Models

DGX agent

arXiv:2606.06188v1 Announce Type: new Abstract: Recent work has sought to understand Large Language Models (LLMs) reasoning, yet a principled, model-intrinsic signal that captures its layer-wise reaso

researcharxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A Latent Variable Framework for Scaling Laws in Large Language Models

DGX agent

arXiv:2512.06553v2 Announce Type: replace-cross Abstract: We propose a statistical framework built on latent variable modeling for scaling laws of large language models (LLMs). Our work is motivated b

researcharxiv-cs-lg
4 Jun 2026
Applications

An Ensembled Latent Factor Model via Differential Evolution and Gradient Descent Optimization

DGX agent

arXiv:2606.04408v1 Announce Type: cross Abstract: High-dimensional and incomplete (HDI) data are prevalent in many real-world big data scenarios. Latent factor models serve as a common representation

applicationsarxiv-cs-ai
4 Jun 2026
Research

Audio Interaction Model

DGX agent

arXiv:2606.05121v1 Announce Type: cross Abstract: Audio is an inherently interactive modality, yet today's Large Audio Language Models (LALMs) are offline, and streaming audio models each handle only

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Discourse-Role Labels as Presentation-Time Variables for Context Use in Language Models

DGX agent

arXiv:2606.04109v1 Announce Type: new Abstract: Context-augmented language model systems often wrap supplied content with labels such as Reference:, Evidence:, Instruction:, Note:, or Example:, but th

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning

DGX agent

arXiv:2606.04986v1 Announce Type: new Abstract: Recent studies have explored Vision-Language Models (VLMs) for food analysis. However, most existing methods rely primarily on supervised fine-tuning (S

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Interfaze: The Future of AI is built on Task-Specific Small Models

DGX agent

arXiv:2602.04101v2 Announce Type: replace Abstract: We present Interfaze, a native hybrid model that fuses task-specific deep neural networks (CNNs and DNNs) directly into a transformer decoder throug

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

LimiX-2M: Mitigating Low-Rank Collapse and Attention Bottlenecks in Tabular Foundation Models

DGX agent

arXiv:2606.04485v1 Announce Type: new Abstract: Tabular foundation models (TFMs) increasingly rival tree ensembles, but their performance is often compute-inefficient: with standard affine scalar toke

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

LoopMoE: Unifying Iterative Computation with Mixture-of-Experts for Language Modeling

DGX agent

arXiv:2606.04438v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) and looped architectures scale models along two orthogonal axes, namely parameter capacity and effective depth. However, main

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

DGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SPLIT-PINN: Separable Probability Learning Technique via Physics-Informed Neural Networks for High-Dimensional Probabilistic Modeling

DGX agent

arXiv:2606.04000v1 Announce Type: cross Abstract: We present a probabilistic modeling framework for incorporating small-scale spatial heterogeneity into macroscopic descriptions of material behavior f

model-releasesarxiv-cs-lg
4 Jun 2026
Applications

Towards Evaluating the Robustness of Visual State Space Models

DGX agent

arXiv:2406.09407v3 Announce Type: replace Abstract: Vision State Space Models (VSSMs), a novel architecture that combines the strengths of recurrent neural networks and latent variable models, have de

applicationsarxiv-cs-cv
4 Jun 2026
Model Releases

transitions like this are why we think it's helpful to have a provider-agnostic harness we used to talk more about swapping models when the …

DGX agent

transitions like this are why we think it's helpful to have a provider-agnostic harness we used to talk more about swapping models when the latest and greatest came out -- but the latest and greatest

model-releasesharrison-chase--x
4 Jun 2026
Model Releases

UniCAD: A Unified Benchmark and Universal Model for Multi-Modal Multi-Task CAD

DGX agent

arXiv:2606.05058v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) underpins modern engineering and manufacturing by enabling the creation of precise, editable 3D models. However, CAD resea

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

VGGSounder: Audio-Visual Evaluations for Foundation Models

DGX agent

arXiv:2508.08237v4 Announce Type: replace-cross Abstract: The emergence of audio-visual foundation models underscores the importance of reliably assessing their multi-modal understanding. The VGGSound

model-releasesarxiv-cs-ai
4 Jun 2026
Tutorials

A Close Look At World Model Recovery In Supervised Fine-Tuned LLM Planners

DGX agent

arXiv:2606.03685v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) improves end-to-end classical planning in large language models (LLMs), but do these models also learn to represent and r

tutorialsarxiv-cs-ai
3 Jun 2026
Tools

[AINews] Microsoft Build: MAI-Thinking-1 and MAI Family models

DGX agent

Microsoft announced the MAI-Thinking-1 and MAI family of models at their Build conference, representing new advances in their AI model lineup. These models likely focus on improved reasoning capabilit

toolslatent-space
3 Jun 2026
Model Releases

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation…

DGX agent

Another banger open-source release. Miso One is an 8B text-to-speech model with real emotional range, so voiceovers carry warmth, hesitation, and excitement instead of sounding flat. It's purpose-buil

model-releasesdair-ai--x
3 Jun 2026
Model Releases

Automated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging

DGX agent

arXiv:2606.02809v1 Announce Type: new Abstract: Evaluating vision-language models (VLMs) on medical images requires benchmarks that are clinically grounded, scalable, and controlled for evaluation con

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

BehaviorBench: Modeling Real-World User Decisions from Behavioral Traces

DGX agent

arXiv:2606.02798v1 Announce Type: new Abstract: Many decision-support settings require systems that adapt to individual users, but evaluation data for this problem remain limited. Existing benchmarks

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Consistent Yet Wrong: Evidence Insensitivity in Spatial Vision-Language Models

DGX agent

arXiv:2606.02742v1 Announce Type: new Abstract: Spatial reasoning is fundamental to robotics, autonomy, and embodied AI, yet modern vision-language models (VLMs) remain unreliable on metric distance q

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Google introduces Gemma 4 12B, a unified, encoder-free open multimodal model that can run locally on devices with 16GB of VRAM or unified memory (Carl Franzen/VentureBeat)

DGX agent

Carl Franzen / VentureBeat: Google introduces Gemma 4 12B, a unified, encoder-free open multimodal model that can run locally on devices with 16GB of VRAM or unified memory — While many AI open source

model-releasestechmeme
3 Jun 2026
Model Releases

ReaLM: Residual Quantization Bridging Knowledge Graph Embeddings and Large Language Models

DGX agent

arXiv:2510.09711v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently emerged as a powerful paradigm for Knowledge Graph Completion (KGC), offering strong reasoning and

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Reasoning Structure of Large Language Models

DGX agent

arXiv:2606.03883v1 Announce Type: new Abstract: Large reasoning models (LRMs) are often evaluated using metrics such as final-answer accuracy or token count. However, identical scores on these metrics

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

RogueMerge: Robust and Unified Attacks against LLM Model Merging

DGX agent

arXiv:2606.03344v1 Announce Type: cross Abstract: Model merging composes specialized capabilities into a single LLM by aggregating task vectors sourced from unverified public platforms, exposing a cri

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

SVHalluc: Benchmarking Speech-Vision Hallucination in Audio-Visual Large Language Models

DGX agent

arXiv:2606.02642v1 Announce Type: cross Abstract: Despite the success of audio-visual large-language models (LLMs), they can produce plausible but ungrounded outputs, termed hallucination. Existing be

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Synthetic Hallucinations, Real Gains: Hard Negatives from Frontier Models for FIM Hallucination Mitigation

DGX agent

arXiv:2606.03130v1 Announce Type: new Abstract: Small open-source code models that power IDE autocomplete still emit hallucinated Fill-in-the-Middle (FIM) completions: syntactically natural calls to m

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models

DGX agent

arXiv:2512.22539v2 Announce Type: replace-cross Abstract: While Vision-Language-Action models (VLAs) are rapidly advancing towards generalist robot policies, it remains difficult to quantitatively und

model-releasesarxiv-cs-cv
3 Jun 2026
Safety

When Models Refuse: Political Steerability and Feature Richness as Measures of Ideological Depth

DGX agent

arXiv:2508.21448v3 Announce Type: replace Abstract: Large language models (LLMs) sometimes refuse to follow benign instructions, such as declining to argue a political position or adopt a stated perso

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

Assessment of Generative Named Entity Recognition in the Era of Large Language Models

DGX agent

arXiv:2601.17898v2 Announce Type: replace Abstract: Named entity recognition (NER) is evolving from a sequence labeling task into a generative paradigm with the rise of large language models (LLMs). W

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Beyond Isolated Behaviors: Hierarchical User Modeling for LLM Personalization

DGX agent

arXiv:2606.02300v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across diverse domains, yet personalizing their outputs to individual users remai

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Business Utility of Large Language Models as Exploratory Data Analysis Agents

DGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translation

DGX agent

arXiv:2606.01048v1 Announce Type: new Abstract: We propose Decoupled Residual Denoising Diffusion models (DRDD) for unified and data-efficient image-to-image (I2I) translation. While diffusion models

researcharxiv-cs-cv
2 Jun 2026
Model Releases

DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?

DGX agent

arXiv:2505.16915v3 Announce Type: replace-cross Abstract: While recent Text-to-Image (T2I) models show impressive capabilities in synthesizing images from brief descriptions, they struggle with the lo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Domain-Shift-Aware Conformal Prediction for Large Language Models

DGX agent

arXiv:2510.05566v2 Announce Type: replace-cross Abstract: Large language models have achieved impressive performance across diverse tasks. However, their tendency to produce overconfident and factuall

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Dynamic Proxy-Mixing: Transferring Replay Controllers from Small to Large Models for Continual Instruction Tuning

DGX agent

arXiv:2606.00400v1 Announce Type: new Abstract: Continual instruction tuning updates a language model through a sequence of new domains, yet each update can progressively erode previously learned capa

model-releasesarxiv-cs-lg
2 Jun 2026
Research

Emergent Transfer of a Physics Foundation Model from Simulation to Laboratory Turbulence

DGX agent

arXiv:2606.01470v1 Announce Type: cross Abstract: Whether physics foundation models can be usefully deployed on laboratory experiments remains an open question for scientific machine learning (ML). We

researcharxiv-cs-ai
2 Jun 2026
Model Releases

From Moments to Models: Graphon-Mixture Learning for Mixup and Contrastive Learning

DGX agent

arXiv:2510.03690v4 Announce Type: replace Abstract: Real-world graph datasets often arise from mixtures of populations, where graphs are generated by multiple distinct underlying distributions. In thi

model-releasesarxiv-cs-lg
2 Jun 2026
Research

GUDA: Counterfactual Group-wise Training Data Attribution for Diffusion Models via Unlearning

DGX agent

arXiv:2601.22651v2 Announce Type: replace-cross Abstract: Training-data attribution for vision generative models aims to identify which training data influenced a given output. While most methods scor

researcharxiv-cs-ai
2 Jun 2026
Research

Learning Action-Conditional and Object-Centric Gaussian Splatting World Models for Rigid Objects

DGX agent

arXiv:2606.01950v1 Announce Type: cross Abstract: World models enable intelligent agents to predict the consequences of their actions on the environment. In this paper, we propose Multi Rigid Object G

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Machine Learning Surrogate Modeling for Homogenization of Hyperelastic Materials with Boolean Microstructures

DGX agent

arXiv:2606.00938v1 Announce Type: cross Abstract: Data-driven surrogate models are an alternative to numerical homogenization of heterogeneous materials. In this contribution, a supervised learning ap

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Margin Adaptive DPO: Leveraging Reward Model for Granular Control in Preference Optimization

DGX agent

arXiv:2510.05342v2 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) has emerged as a simple and effective method for aligning large language models. However, its reliance on

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models

DGX agent

arXiv:2606.01060v1 Announce Type: cross Abstract: Preference alignment has substantially improved the observable behavior of large language models, yet it remains unclear what alignment changes intern

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

PaperVoyager : Building Interactive Web with Visual Language Models

DGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Pramana: Fine-Tuning Large Language Models for Epistemic Reasoning through Navya-Nyaya

DGX agent

arXiv:2604.04937v1 Announce Type: cross Abstract: Large language models produce fluent text but struggle with systematic reasoning, often hallucinating confident but unfounded claims. When Apple resea

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

ProbeScale: Probing Analysis to Optimize Neural Scaling Laws for Efficient Small Language Model Inference

DGX agent

arXiv:2606.01806v1 Announce Type: cross Abstract: Small Language Models (SLMs) offer a balance between capability and computational feasibility. Neural scaling laws inform their optimal training, sugg

model-releasesarxiv-cs-ai
2 Jun 2026
Local Ai

RA-LWLM: Retrieval-Augmented In-Context Localization with Wireless Foundation Models

DGX agent

arXiv:2606.01899v1 Announce Type: cross Abstract: Wireless localization is a fundamental capability of sixth-generation (6G) networks. Conventional model-based methods require accurate modeling of the

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

Revise, Don't Freeze: Sampler-Matched Training for Self-Correcting Masked Diffusion Language Models

DGX agent

arXiv:2606.01026v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) re-predict every position at each denoising step, but standard samplers commit tokens once revealed, leaving th

model-releasesarxiv-cs-cl
2 Jun 2026
← Previous
1…8283848586…1261
Next →