AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
9 Jun 2026

A Machine Learning-Enhanced Hopf-Cole Formulation for Nonlinear Gas Flow in Porous Media

ResearchDGX agent

arXiv:2603.11250v2 Announce Type: replace-cross Abstract: Accurate modeling of gas flow through porous media is critical for many technological applications, including reservoir performance prediction

An Agency-Transferring Model-Free Policy Enhancement Technique

SafetyDGX agent

arXiv:2606.09825v1 Announce Type: cross Abstract: Training reinforcement learning (RL) policies from scratch is costly: it requires careful reward and environment design, extensive tuning, and substan

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs

Model ReleasesDGX agent

arXiv:2606.07643v1 Announce Type: cross Abstract: Recent advances in Omni-Multimodal Large Language Models (Omni-MLLMs) have enabled strong integration of vision, audio, and language. However, their a

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond Pass Rate: A Multilingual, Execution-Grounded Evaluation of Open Code LLMs

Model ReleasesDGX agent

arXiv:2606.08840v1 Announce Type: new Abstract: Code generation models are typically compared using compact execution benchmarks and aggregate pass rates, but such summaries obscure how performance va

CRAG: Can 3D Generative Models Help 3D Assembly?

ResearchDGX agent

arXiv:2602.22629v2 Announce Type: replace Abstract: Most existing 3D assembly methods treat the problem as pure pose estimation, rearranging observed parts via rigid transformations. In contrast, huma

CRANE: Knowledge Editing for Reasoning MLLMs

Model ReleasesDGX agent

arXiv:2606.09033v1 Announce Type: new Abstract: The emergence of reasoning multimodal large language models (MLLMs), which generate explicit chain-of-thought (CoT) reasoning before producing answers,

Decomposable Neuro Symbolic Regression

Model ReleasesDGX agent

arXiv:2511.04124v3 Announce Type: replace Abstract: Symbolic regression (SR) models complex systems by discovering mathematical expressions that capture underlying relationships in observed data. Howe

Defend against frontier cyber models: Cloudflare's architecture as customer zero

IndustryDGX agent

In our post about Project Glasswing, we made the argument that the architecture around a vulnerability matters more than the speed of the patch. Here we walk through what that architecture looks like,

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks

Model ReleasesDGX agent

arXiv:2606.07970v1 Announce Type: cross Abstract: Current open-weight large language models (LLMs) are prone to malicious finetuning attacks, which could compromise the safety alignment of LLMs with o

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

AgentsDGX agent

arXiv:2606.09273v1 Announce Type: new Abstract: 3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane

Google announces Gemini 3.5 Live Translate for instant voice-to-voice translation

Model ReleasesDGX agent

Gemini 3.5 Live Translate is Google's latest audio model delivering near real-time speech-to-speech translation in over 70 languages. The model automatically detects 70+ languages and generates smooth

How Much Capacity Does EEG Denoising Need? Ultra-Compact Networks reveal Benchmark Saturation and Metric-Utility Gap

Model ReleasesDGX agent

arXiv:2606.08594v1 Announce Type: new Abstract: Deep learning EEG denoising architectures have scaled from tens of thousands to tens of millions of parameters, yet no prior study has isolated model ca

IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation

Model ReleasesDGX agent

arXiv:2601.04498v2 Announce Type: replace-cross Abstract: Infographics are composite visual artifacts that combine data visualizations with textual and illustrative elements to communicate information

PEDRA: Evaluating the Realism of Pedestrian Dynamics in Video Generation

Model ReleasesDGX agent

arXiv:2510.20182v2 Announce Type: replace Abstract: Pedestrian simulation traditionally relies on expert-tuned, hand-crafted models that limit scalability and generalization. Meanwhile, large-scale vi

Petri Net Modeling and Deadlock-Free Scheduling of Attachable Heterogeneous AGV Systems

SafetyDGX agent

arXiv:2508.00724v2 Announce Type: replace-cross Abstract: The increasing demand for flexible automation has accelerated the adoption of heterogeneous automated guided vehicles (AGVs). This work invest

PLAGUE: Plug-and-play framework for Lifelong Adaptive Generation of Multi-turn Exploits

Model ReleasesDGX agent

arXiv:2510.17947v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are improving at an exceptional rate. With the advent of agentic workflows, multi-turn dialogue has become the de

RecurGuard: Runtime Monitoring for Reasoning-Token Consumption Attacks

Model ReleasesDGX agent

arXiv:2606.07968v1 Announce Type: cross Abstract: Reasoning-capable large language models can be induced to spend their generation budget on injected decoy tasks rather than answering the user's quest

Scaling Laws for Masked-Reconstruction Transformers on Single-Cell Transcriptomics

Model ReleasesDGX agent

arXiv:2602.15253v2 Announce Type: replace Abstract: Neural scaling laws -- power-law relationships between loss, model size, and data -- have been extensively documented for language and vision transf

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

Model ReleasesDGX agent

arXiv:2606.08034v1 Announce Type: cross Abstract: Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing s

Subtitle-Aligned Fine-Tuning of Whisper for Swiss German ASR: Benchmark Contamination, Convention Mismatch, and an Honest Baseline at 25.6% WER (13.8% cWER)

Model ReleasesDGX agent

arXiv:2606.07608v1 Announce Type: cross Abstract: We present a systematic study of fine-tuning OpenAI's Whisper large-v3 for Swiss German ASR, using 1,367 hours of broadcast speech paired with Standar

TLDR: Compressing Audio Tokens for Efficient Autoregressive Text-to-Speech

ResearchDGX agent

arXiv:2606.09019v1 Announce Type: cross Abstract: Codec-based autoregressive (AR) speech language models have achieved strong text-to-speech (TTS) quality by modeling speech as sequences of discrete a

VATS: Exploiting Implicit Authority in Error-Path Injection via Systematic Mutation

Model ReleasesDGX agent

arXiv:2606.07992v1 Announce Type: new Abstract: As the Model Context Protocol (MCP) standardizes tool-calling for autonomous agents, it introduces a critical, unexamined attack surface: the error-hand

VisualFLIP: Do Predictions Depend on Task-Critical Visual Evidence in Multimodal Reasoning?

Model ReleasesDGX agent

arXiv:2606.07872v1 Announce Type: new Abstract: When a multimodal large language model answers a visual reasoning question correctly, is the prediction actually supported by the task-critical visual e

Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning

SafetyDGX agent

arXiv:2602.01357v2 Announce Type: replace Abstract: Self-play post-training methods has emerged as an effective approach for finetuning large language models and turn the weak language model into stro

Zero-Shot Learning in Industrial Scenarios: New Large-Scale Benchmark, Challenges and Baseline

Model ReleasesDGX agent

arXiv:2606.07965v1 Announce Type: new Abstract: Large Visual Language Models (LVLMs) have achieved remarkable success in vision tasks. However, the significant differences between industrial and natur

ZIPP:Zero-shot Image Personalization from Personas

Model ReleasesDGX agent

arXiv:2606.08841v1 Announce Type: new Abstract: Text-to-image diffusion models are increasingly deployed in open-ended creative contexts, yet their outputs remain impersonal, optimized for aggregate a

8 Jun 2026

CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions

Model ReleasesDGX agent

arXiv:2606.06526v1 Announce Type: new Abstract: Large language models have made substantial progress on mathematical reasoning, but existing benchmarks typically evaluate well-specified problems with

Entropy as a Structural Prior: How a Log-Barrier on DiT Belief Space Drives Musical Diversity and Development

Model ReleasesDGX agent

arXiv:2606.07207v1 Announce Type: cross Abstract: Confidence-based loss weighting is usually avoided in generative models because it accelerates errors when the model is confidently wrong, but this in

Explaining Unsupervised Disease Staging in Huntington's Disease: Insights into Model Representations and Clusters

ResearchDGX agent

arXiv:2606.07135v1 Announce Type: new Abstract: Huntington's disease (HD) is a progressive neurodegenerative disorder that affects motor, cognitive, and behavioral functions, where accurate characteri

LiQSS: Post-Transformer Linear Quantum-Inspired State-Space Tensor Networks for Real-Time 6G

Model ReleasesDGX agent

arXiv:2601.12375v3 Announce Type: replace-cross Abstract: Proactive and agentic control in Sixth-Generation (6G) Open Radio Access Networks (O-RAN) requires control-grade prediction under stringent Ne

Mechanistic Evidence for Faithfulness Decay in Chain-of-Thought Reasoning

TutorialsDGX agent

arXiv:2602.11201v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) explanations are widely used to interpret how language models solve complex problems, yet it remains unclear whether these st

Principles and Practice of Deep Representation Learning: or a Mathematical Theory of Memory

ResearchDGX agent

arXiv:2606.06624v1 Announce Type: new Abstract: In the current era of deep learning and especially generative models, there is significant investment in training very large generative models. Thus far

Quantum-Inspired Trace-Augmented Evidence Selection for Reasoning over Structured Hypothesis Spaces

Model ReleasesDGX agent

arXiv:2606.06941v1 Announce Type: new Abstract: Large language models (LLMs) now solve a wide range of expert-level exams at or above human level, yet remain brittle on specialised, evidence-intensive

SafeGene: Reusable Adapters for Transferable Safety Alignment

SafetyDGX agent

arXiv:2606.06519v1 Announce Type: new Abstract: Open-weight LLMs are increasingly fine-tuned into customized assistants, but downstream fine-tuning can weaken safety alignment and make models more vul

Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors

Model ReleasesDGX agent

arXiv:2606.06891v1 Announce Type: new Abstract: Despite advances in 3D scene understanding, existing 3D Large Multimodal Models operate in offline settings, requiring complete scene observations or pr

UrduMMLU: A Massive Multitask Benchmark for Urdu Language Understanding

Model ReleasesDGX agent

arXiv:2606.07167v1 Announce Type: cross Abstract: Meaningful multilingual evaluation must test models in the target language and educational context. Urdu, spoken by more than 230 million people, lack

WorldBench: A Challenging and Visually Diverse Multimodal Reasoning Benchmark

Model ReleasesDGX agent

arXiv:2606.06538v1 Announce Type: new Abstract: In real-world applications, models are expected to perform reliably across diverse settings. Yet, many existing multimodal benchmarks expand task types

6 Jun 2026

Can AI Refute Economic Theory? Evidence from Beyond the Knowledge Cutoff

Model ReleasesDGX agent

arXiv:2606.05383v1 Announce Type: cross Abstract: Can artificial intelligence (AI) refute economic theory? I document experiments in which I asked several AI models (Gemini, Refine, Claude, and ChatGP

DPBench: Structural Determinants of Multi-Agent LLM Coordination Under Simultaneous Resource Contention

Model ReleasesDGX agent

arXiv:2602.13255v2 Announce Type: replace Abstract: We present DPBench, a benchmark for evaluating coordination in multi-agent systems built from large language models. Existing benchmarks measure tas

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

Model ReleasesDGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

Ideogram 4.0 feels good

Local AiDGX agent

Ideogram 4.0 is a frontier text-to-image foundation model released as an open-weight model with a commercial license. The model delivers frontier-grade text rendering across languages, bounding-box la

Residual Modeling for High-Fidelity Learned Compression of Scientific Data

SafetyDGX agent

arXiv:2606.05389v1 Announce Type: new Abstract: Lossy compression is essential for massive spatiotemporal data from scientific simulations. Learned compressors can achieve high compression ratios at m

Trust, but Don't Verify: Epistemic Blind Spots in LLM Source Evaluation

Model ReleasesDGX agent

arXiv:2606.05403v1 Announce Type: cross Abstract: Language models increasingly act as epistemic proxies, synthesizing evidence from multiple sources to inform decisions. Whether they evaluate the qual

5 Jun 2026

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models

SafetyDGX agent

arXiv:2602.12628v4 Announce Type: replace Abstract: Simulation offers a scalable and low-cost way to enrich vision-language-action (VLA) training, reducing reliance on expensive real-robot demonstrati

Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.21288v2 Announce Type: replace-cross Abstract: Autonomous driving is an important and safety-critical task, and recent advances in LLMs/VLMs have opened new possibilities for reasoning and

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

Model ReleasesDGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

Noise-Aware Visual Representation Learning for Medical Visual Question Answering

Model ReleasesDGX agent

arXiv:2606.05535v1 Announce Type: new Abstract: Medical visual question answering (Med-VQA) has strong potential for clinical decision support by enabling AI models to interpret medical images and ans

The latest AI news we announced in May 2026

Model ReleasesDGX agent

Google's May 2026 AI updates center on the new 'agentic' era, featuring the Gemini 3.5 model and Gemini Omni for advanced reasoning and creation. Gemini Omni is a new model that can create anything fr

4 Jun 2026

AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms

Model ReleasesDGX agent

arXiv:2602.09464v2 Announce Type: replace-cross Abstract: Vericoding refers to the generation of formally verified code from rigorous specifications. Recent AI models show promise in vericoding, but a

Aligning Deep Implicit Preferences by Learning to Reason Defensively

Model ReleasesDGX agent

arXiv:2510.11194v3 Announce Type: replace Abstract: Personalized alignment is crucial for enabling Large Language Models (LLMs) to engage effectively in user-centric interactions. However, current met

CADET: A Modular Platform for Evaluating Distributed Cooperative Autonomy in Connected Autonomous Vehicles

Model ReleasesDGX agent

arXiv:2606.04072v1 Announce Type: cross Abstract: Deep learning models are increasingly central to autonomous vehicle (AV) pipelines, yet their integration has traditionally followed a monolithic desi

Caliper: Probing Lexical Anchors versus Causal Structure in LLMs

Model ReleasesDGX agent

arXiv:2606.04915v1 Announce Type: new Abstract: Large language models reach 50 to 70% accuracy on causal reasoning benchmarks such as CLadder, but it is unclear whether this reflects structural reason

Depth-Attention: Cross-Layer Value Mixing for Language Models

ResearchDGX agent

arXiv:2606.05014v1 Announce Type: new Abstract: Self-attention selects information freely across the sequence, but across depth, Transformers merely add each layer's output to the residual stream, so

dMX: Differentiable Mixed-Precision Assignment for Low-Precision Floating-Point Formats

Model ReleasesDGX agent

arXiv:2606.04115v1 Announce Type: cross Abstract: Quantizing large language models (LLMs) to low-precision floating-point representations is central to efficient deployment, yet applying a single bit-

FindIt: A Format-Informed Visual Detection Benchmark for Generalist Multimodal LLMs

Model ReleasesDGX agent

arXiv:2606.04282v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are predominantly evaluated on free-form vision-language tasks such as visual question answering, captioning, a

Gravity-Aware Hierarchical Routing for Lightweight SensorLLM on Human Activity Recognition

Model ReleasesDGX agent

arXiv:2606.04019v1 Announce Type: cross Abstract: Recent studies on sensor-language alignment have shown that two-stage frameworks can improve the semantic modeling ability of wearable-sensor human ac

https://ollama.com/library/gemma4

Local AiDGX agent

Gemma4 is a language model available through Ollama's model library that users can download and run locally. The entry likely provides information about the model's specifications, capabilities, and h

Hybrid Adversarial Defence for Natural Language Understanding Tasks

SafetyDGX agent

arXiv:2606.04612v1 Announce Type: new Abstract: Large Language Models (LLMs) are vulnerable both to hallucination and adversarial manipulation. Although these problems are closely related, existing de

Literature-Guided Minimax Optimization of Virtual Epilepsy Neurostimulation

Model ReleasesDGX agent

arXiv:2606.04339v1 Announce Type: new Abstract: Computational models of epilepsy promise patient-specific treatment design, but most optimization workflows still search for parameters that perform wel

Nemotron 3 Ultra now available on AI Gateway

Model ReleasesDGX agent

Nemotron 3 Ultra, NVIDIA's advanced language model, is now accessible through Vercel's AI Gateway, enabling developers to integrate this model into their applications alongside other LLM options. The

← Previous
1…278279280281282…1034
Next →