AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,083
  • Agents7,615
  • Applications5,445
  • Concepts5
  • Hardware1,866
  • Industry6,184
  • Local Ai4,979
  • Model Releases24,164
  • Research20,258
  • Safety13,457
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
89,083Total entries
1Added by human
89,082Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,195 results
Model Releases

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

DGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

model-releasesarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech

DGX agent

arXiv:2606.09019v1 Announce Type: cross Abstract: Codec-based autoregressive (AR) speech language models have achieved strong text-to-speech (TTS) quality by modeling speech as sequences of discrete a

researcharxiv-cs-ai
9 Jun 2026
Model Releases

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

DGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

DGX agent

arXiv:2606.07872v1 Announce Type: new Abstract: When a multimodal large language model answers a visual reasoning question correctly, is the prediction actually supported by the task-critical visual e

model-releasesarxiv-cs-cv
9 Jun 2026
Safety

Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning

DGX agent

arXiv:2602.01357v2 Announce Type: replace Abstract: Self-play post-training methods has emerged as an effective approach for finetuning large language models and turn the weak language model into stro

safetyarxiv-cs-lg
9 Jun 2026
Model Releases

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

DGX agent

arXiv:2606.07965v1 Announce Type: new Abstract: Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natur

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ZIPP:Zero-shot Image Personalization from Personas

DGX agent

arXiv:2606.08841v1 Announce Type: new Abstract: Text-to-image diffusion models are increasingly deployed in open-ended creative contexts, yet their outputs remain impersonal, optimized for aggregate a

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions

DGX agent

arXiv:2606.06526v1 Announce Type: new Abstract: Large language models have made substantial progress on mathematical reasoning, but existing benchmarks typically evaluate well-specified problems with

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development

DGX agent

arXiv:2606.07207v1 Announce Type: cross Abstract: Confidence-based loss weighting is usually avoided in generative models because it accelerates errors when the model is confidently wrong, but this in

model-releasesarxiv-cs-lg
8 Jun 2026
Research

Explaining Unsupervised Disease Staging in Huntington's Disease: Insights into Model Representations and Clusters

DGX agent

arXiv:2606.07135v1 Announce Type: new Abstract: Huntington's disease (HD) is a progressive neurodegenerative disorder that affects motor, cognitive, and behavioral functions, where accurate characteri

researcharxiv-cs-lg
8 Jun 2026
Model Releases

LiQSS: Post-Transformer Linear Quantum-Inspired State-Space Tensor Networks for Real-Time 6G

DGX agent

arXiv:2601.12375v3 Announce Type: replace-cross Abstract: Proactive and agentic control in Sixth-Generation (6G) Open Radio Access Networks (O-RAN) requires control-grade prediction under stringent Ne

model-releasesarxiv-cs-lg
8 Jun 2026
Tutorials

Mechanistic Evidence for Faithfulness Decay in Chain-of-Thought Reasoning

DGX agent

arXiv:2602.11201v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) explanations are widely used to interpret how language models solve complex problems, yet it remains unclear whether these st

tutorialsarxiv-cs-cl
8 Jun 2026
Research

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory

DGX agent

arXiv:2606.06624v1 Announce Type: new Abstract: In the current era of deep learning and especially generative models, there is significant investment in training very large generative models. Thus far

researcharxiv-cs-lg
8 Jun 2026
Model Releases

Quantum-Inspired Trace-Augmented Evidence Selection for Reasoning over Structured Hypothesis Spaces

DGX agent

arXiv:2606.06941v1 Announce Type: new Abstract: Large language models (LLMs) now solve a wide range of expert-level exams at or above human level, yet remain brittle on specialised, evidence-intensive

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

SafeGene: Reusable Adapters for Transferable Safety Alignment

DGX agent

arXiv:2606.06519v1 Announce Type: new Abstract: Open-weight LLMs are increasingly fine-tuned into customized assistants, but downstream fine-tuning can weaken safety alignment and make models more vul

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

DGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

UrduMMLU: A Massive Multitask Benchmark for Urdu Language Understanding

DGX agent

arXiv:2606.07167v1 Announce Type: cross Abstract: Meaningful multilingual evaluation must test models in the target language and educational context. Urdu, spoken by more than 230 million people, lack

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

WorldBench: A Challenging and Visually Diverse Multimodal Reasoning Benchmark

DGX agent

arXiv:2606.06538v1 Announce Type: new Abstract: In real-world applications, models are expected to perform reliably across diverse settings. Yet, many existing multimodal benchmarks expand task types

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff

DGX agent

arXiv:2606.05383v1 Announce Type: cross Abstract: Can artificial intelligence (AI) refute economic theory? I document experiments in which I asked several AI models (Gemini, Refine, Claude, and ChatGP

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DPBench: Structural Determinants of Multi-Agent LLM Coordination Under Simultaneous Resource Contention

DGX agent

arXiv:2602.13255v2 Announce Type: replace Abstract: We present DPBench, a benchmark for evaluating coordination in multi-agent systems built from large language models. Existing benchmarks measure tas

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

DGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

model-releasesarxiv-cs-ai
6 Jun 2026
Local Ai

Ideogram 4.0 feels good

DGX agent

Ideogram 4.0 is a frontier text-to-image foundation model released as an open-weight model with a commercial license. The model delivers frontier-grade text rendering across languages, bounding-box la

local-air-stablediffusion
6 Jun 2026
Safety

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

DGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Trust, but Don't Verify: Epistemic Blind Spots in LLM Source Evaluation

DGX agent

arXiv:2606.05403v1 Announce Type: cross Abstract: Language models increasingly act as epistemic proxies, synthesizing evidence from multiple sources to inform decisions. Whether they evaluate the qual

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

DGX agent

arXiv:2602.12628v4 Announce Type: replace Abstract: Simulation offers a scalable and low-cost way to enrich vision-language-action (VLA) training, reducing reliance on expensive real-robot demonstrati

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving

DGX agent

arXiv:2601.21288v2 Announce Type: replace-cross Abstract: Autonomous driving is an important and safety-critical task, and recent advances in LLMs/VLMs have opened new possibilities for reasoning and

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

DGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

Noise-Aware Visual Representation Learning for Medical Visual Question Answering

DGX agent

arXiv:2606.05535v1 Announce Type: new Abstract: Medical visual question answering (Med-VQA) has strong potential for clinical decision support by enabling AI models to interpret medical images and ans

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

The latest AI news we announced in May 2026

DGX agent

Google's May 2026 AI updates center on the new 'agentic' era, featuring the Gemini 3.5 model and Gemini Omni for advanced reasoning and creation. Gemini Omni is a new model that can create anything fr

model-releasesgoogle-ai
5 Jun 2026
Model Releases

AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms

DGX agent

arXiv:2602.09464v2 Announce Type: replace-cross Abstract: Vericoding refers to the generation of formally verified code from rigorous specifications. Recent AI models show promise in vericoding, but a

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Aligning Deep Implicit Preferences by Learning to Reason Defensively

DGX agent

arXiv:2510.11194v3 Announce Type: replace Abstract: Personalized alignment is crucial for enabling Large Language Models (LLMs) to engage effectively in user-centric interactions. However, current met

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

CADET: A Modular Platform for Evaluating Distributed Cooperative Autonomy in Connected Autonomous Vehicles

DGX agent

arXiv:2606.04072v1 Announce Type: cross Abstract: Deep learning models are increasingly central to autonomous vehicle (AV) pipelines, yet their integration has traditionally followed a monolithic desi

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Caliper: Probing Lexical Anchors versus Causal Structure in LLMs

DGX agent

arXiv:2606.04915v1 Announce Type: new Abstract: Large language models reach 50 to 70% accuracy on causal reasoning benchmarks such as CLadder, but it is unclear whether this reflects structural reason

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Depth-Attention: Cross-Layer Value Mixing for Language Models

DGX agent

arXiv:2606.05014v1 Announce Type: new Abstract: Self-attention selects information freely across the sequence, but across depth, Transformers merely add each layer's output to the residual stream, so

researcharxiv-cs-cl
4 Jun 2026
Model Releases

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

DGX agent

arXiv:2606.04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

FindIt: A Format-Informed Visual Detection Benchmark for Generalist Multimodal LLMs

DGX agent

arXiv:2606.04282v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are predominantly evaluated on free-form vision-language tasks such as visual question answering, captioning, a

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Gravity-Aware Hierarchical Routing for Lightweight SensorLLM on Human Activity Recognition

DGX agent

arXiv:2606.04019v1 Announce Type: cross Abstract: Recent studies on sensor-language alignment have shown that two-stage frameworks can improve the semantic modeling ability of wearable-sensor human ac

model-releasesarxiv-cs-ai
4 Jun 2026
Local Ai

https://ollama.com/library/gemma4

DGX agent

Gemma4 is a language model available through Ollama's model library that users can download and run locally. The entry likely provides information about the model's specifications, capabilities, and h

local-aiollama--x
4 Jun 2026
Safety

Hybrid Adversarial Defence for Natural Language Understanding Tasks

DGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

Literature-Guided Minimax Optimization of Virtual Epilepsy Neurostimulation

DGX agent

arXiv:2606.04339v1 Announce Type: new Abstract: Computational models of epilepsy promise patient-specific treatment design, but most optimization workflows still search for parameters that perform wel

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Nemotron 3 Ultra now available on AI Gateway

DGX agent

Nemotron 3 Ultra, NVIDIA's advanced language model, is now accessible through Vercel's AI Gateway, enabling developers to integrate this model into their applications alongside other LLM options. The

model-releasesvercel-blog
4 Jun 2026
Model Releases

OckBench: Measuring the Efficiency of LLM Reasoning

DGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Prediction Under Imperfect Compression: A Theory of Approximate MDL

DGX agent

arXiv:2606.04834v1 Announce Type: new Abstract: Minimum Description Length (MDL) formalizes the principle of Occam's razor by optimizing the total description length: L(model)+L(data | model). For seq

researcharxiv-cs-lg
4 Jun 2026
Model Releases

RIDE: An Open Dataset and Benchmark for Train Delay Prediction

DGX agent

arXiv:2606.05070v1 Announce Type: new Abstract: Train delay prediction is an important problem for both passengers and railway operators, yet progress in the field remains difficult to assess due to t

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety

DGX agent

arXiv:2603.10044v2 Announce Type: replace-cross Abstract: A safety score earned on a benchmark need not predict how the same model behaves once it is wrapped in an agentic scaffold the benchmark never

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Stepwise Reasoning Enhancement for LLMs via External Subgraph Generation

DGX agent

arXiv:2606.04454v1 Announce Type: new Abstract: Large language models have shown strong performance in natural language generation and downstream reasoning tasks, but they still struggle with logical

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbations

DGX agent

arXiv:2606.05165v1 Announce Type: cross Abstract: Training Data Attribution (TDA) seeks to trace a model's predictions back to its training data. The gold standard for TDA relies on causal interventio

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

SymTRELLIS: Symmetry-Enforced Voxel Latents for 3D Generation

DGX agent

arXiv:2606.04108v1 Announce Type: cross Abstract: Single-view 3D generative models have achieved impressive visual quality, yet they are not designed to satisfy structural or functional requirements,

model-releasesarxiv-cs-ai
4 Jun 2026
← Previous
1…362363364365366…1338
Next →