AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
89,083Total entries
1Added by human
89,082Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,195 results
Model Releases

Riemannian Optimization for Hadamard Products of Low-Rank Matrices

DGX agent

arXiv:2606.01216v1 Announce Type: new Abstract: The elementwise Hadamard product of two low-rank matrices provides a parameter-efficient model for data with multiplicative structure, but its modeling

model-releasesarxiv-cs-lg
2 Jun 2026
Local Ai
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Structure Enables Effective Self-Localization of Errors in LLMs

DGX agent

arXiv:2602.02416v2 Announce Type: replace Abstract: Self-correction in language models remains elusive. In this work, we explore whether language models can explicitly localize errors in incorrect rea

local-aiarxiv-cs-ai
2 Jun 2026
Research

Subliminal Learning Is Steering Vector Distillation

DGX agent

arXiv:2606.00995v1 Announce Type: new Abstract: Subliminal learning refers to a student language model acquiring a teacher's traits (e.g. a system-prompted preference for owls) when fine-tuned on the

researcharxiv-cs-ai
2 Jun 2026
Research

The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge

DGX agent

arXiv:2606.00829v1 Announce Type: new Abstract: EgoCross evaluates multimodal large language models on egocentric video question answering under substantial domain shift, where test videos come from s

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Toward accurate RUL and SoH estimation using reinforced graph-based physics-informed neural networks enhanced with dynamic weights

DGX agent

arXiv:2507.09766v2 Announce Type: replace-cross Abstract: Accurate estimation of Remaining Useful Life (RUL) and State of Health (SoH) is essential for reliable Prognostics and Health Management (PHM)

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TravelEval: A Comprehensive Benchmarking Framework for Evaluating LLM-Powered Travel Planning Agents

DGX agent

arXiv:2606.01046v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has significantly improved travel planning applications, yet evaluating such models is limited by existi

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages

DGX agent

arXiv:2606.01322v1 Announce Type: cross Abstract: Safety evaluation of Large Language Models (LLMs) remains heavily English-centric, leaving Low-Resource Languages (LRLs), particularly African ones, c

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Variational Learning for Insertion-based Generation

DGX agent

arXiv:2606.02133v1 Announce Type: cross Abstract: Non-monotonic sequence generation methods, such as masked diffusion models, provide a flexible alternative to left-to-right autoregressive modeling by

researcharxiv-cs-ai
2 Jun 2026
Research

When Do Attention Circuits Form? Developmental Trajectories of Capability and Attention-Sink Emergence Across Three 1B-ClassArchitectures

DGX agent

arXiv:2606.02378v1 Announce Type: cross Abstract: We track the developmental trajectory of attention-head circuit formation across three 1B-class language models spanning two architecture families (de

researcharxiv-cs-ai
2 Jun 2026
Model Releases

3DAE: Binaural Quality Assessment for Audio Novel View Synthesis with Spatial Maps and Benchmark

DGX agent

arXiv:2605.30469v1 Announce Type: cross Abstract: 3D audio and novel-view acoustic synthesis models are usually evaluated with global metrics.However, global metrics often hide where and why binaural

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Anchoring LLM Gender Bias to Human Baselines: A Cross-Lingual Audit

DGX agent

arXiv:2605.30804v1 Announce Type: new Abstract: We audit six large language models (LLMs) for gender stereotyping across English, Korean, Chinese, and Japanese. Three were developed primarily for Engl

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Automating Formal Verification with Reinforcement Learning and Recursive Inference

DGX agent

arXiv:2605.30914v1 Announce Type: new Abstract: Automated formal verification remains challenging for large language models because data for proof assistants and verification-aware languages is scarce

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Can LLM Teams Play What? Where? When?

DGX agent

arXiv:2605.30459v1 Announce Type: new Abstract: Large language models (LLMs) remain limited on tasks requiring indirect reasoning, cultural knowledge, and coordinated hypothesis testing. We investigat

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Design and Evaluation of Multi-Agent AI Oracle Systems for Prediction Market Resolution

DGX agent

arXiv:2605.30802v1 Announce Type: cross Abstract: Prediction markets aggregate collective intelligence to forecast uncertain events, but their utility depends on reliable outcome resolution. Existing

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

Distilling Counterfactual Reasoning from Language to Vision: Causal Graph Guided Post-Training for Video Understanding

DGX agent

arXiv:2511.19923v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) have recently shown significant advancements in video understanding, especially in feature alignment, event reas

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution

DGX agent

arXiv:2605.30431v1 Announce Type: new Abstract: Recent progress in video diffusion models has enabled remarkable generative fidelity, yet leveraging these priors for restoration remains limited by the

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

FAM-Bench: A Multimodal Benchmark for Condition-Aware Food-as-Medicine Reasoning

DGX agent

arXiv:2605.31410v1 Announce Type: new Abstract: Food-as-Medicine requires models to reason beyond what a dish is or what nutrition it contains: they must decide whether a concrete food choice is appro

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Forecasting with Hyper-Trees

DGX agent

arXiv:2405.07836v5 Announce Type: replace Abstract: We introduce Hyper-Trees as a novel framework for modeling time series data using gradient boosted trees. Unlike conventional tree-based approaches

safetyarxiv-cs-lg
1 Jun 2026
Research

Generative Models and Statistical Validation

DGX agent

arXiv:2605.30453v1 Announce Type: cross Abstract: Generative machine learning has become an essential tool in theoretical and experimental physics, especially in the context of fast surrogates and den

researcharxiv-cs-lg
1 Jun 2026
Research

Graph Machine Learning in the Era of Large Language Models (LLMs)

DGX agent

arXiv:2404.14928v3 Announce Type: replace-cross Abstract: Graphs play an important role in representing complex relationships in various domains like social networks, knowledge graphs, and molecular d

researcharxiv-cs-ai
1 Jun 2026
Industry

Microsoft to unveil new AI models and Windows improvements at Build

DGX agent

Microsoft is heading to San Francisco this week in a bid to win back developers at its Build conference. I've been attending Build since the days when Microsoft called it the Professional Developers C

industrythe-verge-ai
1 Jun 2026
Research

Modeling Covariate Transition for Efficient Estimation of Longitudinal Treatment Effects in Randomized Experiments

DGX agent

arXiv:2605.31443v1 Announce Type: cross Abstract: We present a regression-adjustment framework designed for the estimation of longitudinal treatment effects in randomized experiments under static regi

researcharxiv-cs-lg
1 Jun 2026
Model Releases

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820…

DGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

model-releasesclem-delangue--x
1 Jun 2026
Research

Target-Side Paraphrase Augmentation for Sign Language Translation with Large Language Models

DGX agent

arXiv:2605.31393v1 Announce Type: cross Abstract: Sign language translation (SLT) remains constrained by limited paired sign-video/text corpora and heavy-tailed target vocabularies. We study target-si

researcharxiv-cs-ai
1 Jun 2026
Model Releases

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

DGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

model-releasesgoogle-cloud-ai
1 Jun 2026
Model Releases

Thinking in Structures: Evaluating Spatial Intelligence in Constraint-Governed Spaces

DGX agent

arXiv:2602.07864v2 Announce Type: replace Abstract: Spatial intelligence is crucial for vision--language models (VLMs), yet many scene-centric benchmarks evaluate unconstrained environments where a si

model-releasesarxiv-cs-cv
1 Jun 2026
Hardware

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how t…

DGX agent

This pod was an incredible gift to the community: not only our first pod about @xAI, but Ethan really indulged on all our questions on how to train a SOTA Videogen world model, including specific area

hardwareswyx--x
1 Jun 2026
Model Releases

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness

DGX agent

arXiv:2605.30911v1 Announce Type: cross Abstract: Hallucination remains one of the key challenges undermining the reliability of Large Vision-Language Models (LVLMs). But what makes an LVLM hallucinat

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

XLGoBench: Detecting cross-lingual skill gaps with algorithmic tasks

DGX agent

arXiv:2605.30788v1 Announce Type: cross Abstract: We introduce a set of synthetic algorithmic tasks to detect cross-lingual gaps in the abilities of large language models. Our benchmark is commensurat

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

11 demos of Gemini Omni and Gemini 3.5 in action

DGX agent

Google unveiled Gemini 3.5 Flash, the latest model combining frontier intelligence with action, and Gemini Omni, a new model that can create anything from any input starting with video. With Omni, use

model-releasesgoogle-ai
29 May 2026
Applications

Accommodation Goes Both Ways: Studying Linguistic Convergence Between Humans and Language Models

DGX agent

arXiv:2605.29278v1 Announce Type: new Abstract: As LLMs become increasingly integrated into daily life, understanding how their presence will shape human linguistic behavior is an open question. We pr

applicationsarxiv-cs-cl
29 May 2026
Model Releases

Architecture-Sensitive Supervised Fine-Tuning for Screen-Conditioned Action Prediction: A PiSAR Benchmark

DGX agent

arXiv:2605.29400v1 Announce Type: new Abstract: We benchmark three supervised fine-tuned models against frontier zero-shot baselines on a 661-row held-out slice of PiSAR (Persona, intent, Screen, Acti

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking Single-Factor Physical Video-to-Audio Generation

DGX agent

arXiv:2605.30339v1 Announce Type: new Abstract: Generative video-to-audio (V2A) models produce highly plausible soundtracks, but it remains unclear whether they capture the underlying physical process

model-releasesarxiv-cs-cv
29 May 2026
Safety

Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling

DGX agent

arXiv:2605.29697v1 Announce Type: new Abstract: In Agentic Search, trajectory-level outcome rewards fail to quantify the behavioral contributions of individual steps, while existing step-level reward

safetyarxiv-cs-ai
29 May 2026
Research

Boosting Image Quality Assessment Performance: Unsupervised Score Fusion by Deep Maximum a Posteriori Estimation

DGX agent

arXiv:2605.30269v1 Announce Type: new Abstract: Over the past decades, numerous Image Quality Assessment (IQA) models have emerged, aiming to predict the perceptual quality of images. However, individ

researcharxiv-cs-cv
29 May 2026
Model Releases

Certified Causal Defense with Generalizable Robustness

DGX agent

arXiv:2408.15451v3 Announce Type: replace Abstract: While machine learning models have proven effective across various scenarios, it is widely acknowledged that many models are vulnerable to adversari

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Cloud CISO Perspectives: How to build an AI-ready security program for the public sector

DGX agent

Welcome to the second Cloud CISO Perspectives for May 2026. Today, Usman Chaudhary, Field CISO, Google Public Sector, offers a guide for CISOs protecting government agencies and critical infrastructur

model-releasesgoogle-cloud-ai
29 May 2026
Hardware

DFlash: Block Diffusion for Flash Speculative Decoding

DGX agent

arXiv:2602.06036v2 Announce Type: replace Abstract: Autoregressive large language models (LLMs) deliver strong performance but require inherently sequential decoding, leading to high inference latency

hardwarearxiv-cs-cl
29 May 2026
Model Releases

EarthShift: a benchmark for measuring robustness to real-world distribution shifts in Earth observation

DGX agent

arXiv:2605.29330v1 Announce Type: new Abstract: Current Earth observation benchmarks focus on measuring performance on diverse tasks and applications, typically measuring generalization in-distributio

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

FinGuard: Detecting Financial Regulatory Non-Compliance in LLM Interactions

DGX agent

arXiv:2605.29427v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed in financial services, a single non-compliant interaction can expose institutions to regulator

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

GPIC: A Giant Permissive Image Corpus for Visual Generation

DGX agent

arXiv:2605.30341v1 Announce Type: cross Abstract: Studying scalable methods for visual generative modeling requires large, accessible, and stable datasets. We introduce GPIC, a Giant Permissive Image

model-releasesarxiv-cs-ai
29 May 2026
Research

HoliTok:A Coutinuous Holistic Tokenization with Robust Dual Capabilities of Speech Generation and Understanding

DGX agent

arXiv:2605.29948v1 Announce Type: cross Abstract: Unified speech foundation models require a holistic tokenization space that is both learnable by language models and decodable into high-quality wavef

researcharxiv-cs-ai
29 May 2026
Model Releases

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

DGX agent

arXiv:2605.29523v1 Announce Type: new Abstract: Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barri

model-releasesarxiv-cs-lg
29 May 2026
Research

KLAS: Using Similarity to Stitch Neural Networks for Improved Accuracy-Efficiency Tradeoffs

DGX agent

arXiv:2605.29259v1 Announce Type: cross Abstract: Given the wide range of deployment targets, flexible model selection is essential for optimizing performance within a given compute budget. Recent wor

researcharxiv-cs-ai
29 May 2026
Model Releases

Less is Enough: Synthesizing Diverse Data in LLM Feature Space with Sparse Autoencoders

DGX agent

arXiv:2602.10388v3 Announce Type: replace-cross Abstract: The diversity of post-training data is critical for effective downstream performance in large language models (LLMs). Many existing approaches

model-releasesarxiv-cs-ai
29 May 2026
Research

MiAD: Mirage Atom Diffusion for De Novo Crystal Generation

DGX agent

arXiv:2511.14426v2 Announce Type: replace-cross Abstract: In recent years, diffusion-based models have demonstrated exceptional performance in searching for simultaneously stable, unique, and novel (S

researcharxiv-cs-ai
29 May 2026
Model Releases

Mind Your Tone: Does Tone Alter LLM Performance?

DGX agent

arXiv:2605.29027v1 Announce Type: new Abstract: The use of Large Language Models (LLMs) is proliferating, yet their performance is observed to vary based on prompting styles and tones. In this study,

model-releasesarxiv-cs-ai
29 May 2026
Research

Neural Scaling Laws for Jet Generation

DGX agent

arXiv:2605.28940v1 Announce Type: cross Abstract: Recently observed empirical scaling laws describe the performance of foundation-type models as three independent key quantities -- dataset size, compu

researcharxiv-cs-lg
29 May 2026
← Previous
1…364365366367368…1338
Next →