AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,514 results
Model Releases

A Fair Benchmarking of Deep Relational Database Learning Models

DGX agent

arXiv:2607.03659v1 Announce Type: cross Abstract: Relational databases (RDBs) are the primary data infrastructure in many enterprises, yet recent deep learning methods designed for RDBs have been eval

model-releasesarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A Technical Survey of Reinforcement Learning Techniques for Large Language Models

DGX agent

arXiv:2507.04136v2 Announce Type: replace Abstract: This survey offers a comprehensive foundation on the integration of RL with language models, highlighting prominent algorithms such as Proximal Poli

model-releasesarxiv-cs-ai
7 Jul 2026
Research

A Unified Framework for In-Context Learning with Causal and Masked Language Models

DGX agent

arXiv:2607.04081v1 Announce Type: new Abstract: In-context learning (ICL) has emerged as a central capability of pretrained language models, yet its theoretical analysis has focused primarily on causa

researcharxiv-cs-lg
7 Jul 2026
Model Releases

Attention Dynamics in Diffusion Models: A Visual Analytics Framework for Human-AI Collaboration

DGX agent

arXiv:2607.02563v1 Announce Type: cross Abstract: Diffusion-based text-to-image models can synthesize complex and highly structured visual content, yet the emergence and evolution of semantic structur

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Banger compression paper from NVIDIA. (bookmark it) Bigger MoE models keep winning on quality, but serving them at interactive latency is st…

DGX agent

Banger compression paper from NVIDIA. (bookmark it) Bigger MoE models keep winning on quality, but serving them at interactive latency is still hard. NVIDIA compresses the hybrid MoE Nemotron-3-Super

model-releasesdair-ai--x
7 Jul 2026
Model Releases

CL-Anomaly: Layer-Adaptive Mixture-of-Experts with Multimodal Large Language Model for Continual Learning in Anomaly Detection

DGX agent

arXiv:2607.02930v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel in diverse vision tasks, but full-parameter retraining is computationally expensive as real-world knowled

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Do Medical Vision Language Models Actually See? A Counterfactual Grounding Framework and Hard-Negative Contrastive Training for Visually-Reliant Medical VLMs

DGX agent

arXiv:2607.03647v1 Announce Type: new Abstract: Large vision language models (VLMs) report strong accuracy on medical question-answering, yet it remains unclear whether they reason from visual evidenc

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

DREAMSTEER: Latent World Models Can Steer VLA Policies During Deployment Without Any Finetuning

DGX agent

arXiv:2607.02865v1 Announce Type: new Abstract: Pretrained vision-language-action (VLA) policies show promising zero-shot generalization, but often fail under deployment-time distribution shift, leadi

model-releasesarxiv-cs-ro
7 Jul 2026
Research

DynaVieW: Schema-Guided World Modeling for Understanding Hierarchical Visual Dynamics

DGX agent

arXiv:2607.04112v1 Announce Type: cross Abstract: Multimodal LLMs struggle to systematically model the temporal evolution of visual scenes in videos or multi-image sequences. Such inputs require model

researcharxiv-cs-ai
7 Jul 2026
Safety

Enhancing the Forecasting Capability of Multi-Model Blending Algorithms for Extreme Precipitation via Joint Use of Station and Gridded Observations

DGX agent

arXiv:2607.04862v1 Announce Type: new Abstract: Accurate extreme precipitation forecasting is critical for disaster mitigation but remains challenging for numerical weather prediction (NWP) models due

safetyarxiv-cs-lg
7 Jul 2026
Model Releases

Erasing Without Collateral Damage: Precise Concept Removal in Diffusion Models

DGX agent

arXiv:2607.05274v1 Announce Type: new Abstract: Training-free concept erasure is an attractive mechanism for controlling text-to-image diffusion models, but precise erasure often comes at the cost of

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Excellent investigation of the mechanics of language models. 🫢A notable blindspot, however: several teams have actually been **directly** c…

DGX agent

Excellent investigation of the mechanics of language models. 🫢A notable blindspot, however: several teams have actually been **directly** comparing the working of LLMs to those of the human brain 🧠 fo

model-releasesyann-lecun--x
7 Jul 2026
Research

Exploring the Rashomon Set for Concept-Based Models

DGX agent

arXiv:2511.19636v2 Announce Type: replace-cross Abstract: In many machine learning problems, there may exist multiple models that achieve nearly identical predictive performance while relying on funda

researcharxiv-cs-ai
7 Jul 2026
Local Ai

Framework for Grouping Local Process Models

DGX agent

arXiv:2607.04856v1 Announce Type: new Abstract: Local Process Models (LPMs) are an underexplored concept in process mining. LPMs describe patterns in event data considering sequence, choice, concurren

local-aiarxiv-cs-lg
7 Jul 2026
Research

Geographic Diversity Beats Data Volume for Cross-Domain Generalization in Zero-Label JEPA Driving World Models

DGX agent

arXiv:2607.04500v1 Announce Type: new Abstract: Self-supervised latent world models can assign a surprise score to driving scenarios without any human labels. A natural follow-up question is whether s

researcharxiv-cs-cv
7 Jul 2026
Model Releases

Geometry of Ordinal Representations in Language Models

DGX agent

arXiv:2607.04167v1 Announce Type: new Abstract: Recent work showed that language models represent character counts on curved 1D manifolds, with attention heads performing geometric transformations to

model-releasesarxiv-cs-lg
7 Jul 2026
Safety

How Utilitarian Are OpenAI's Models Really? Replicating and Reinterpreting Pfeffer, Krugel, and Uhl (2025)

DGX agent

arXiv:2603.22730v2 Announce Type: replace Abstract: Pfeffer, Krugel, and Uhl (2025) report that OpenAI's reasoning model o1-mini produces more utilitarian responses to the trolley problem and footbrid

safetyarxiv-cs-cl
7 Jul 2026
Model Releases

Integrating Neural Encoders in Bayesian Generalized Linear Mixed Models for Multimodal Data

DGX agent

arXiv:2607.04647v1 Announce Type: cross Abstract: Scalable Bayesian inference for generalized linear mixed models (GLMMs) provides uncertainty-aware analysis of correlated longitudinal data, but exist

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Modular Foundation Models for Time-Series Perception in Digital Twins

DGX agent

arXiv:2607.03585v1 Announce Type: new Abstract: Engineering Digital Twins and Prognostics and Health Management (PHM) systems rely on robust perception modules to extract actionable information from h

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models

DGX agent

arXiv:2607.03050v1 Announce Type: cross Abstract: Omni modal large language models (OmniLLMs) have attracted wide attention for their ability to jointly process audio and video, but they generate larg

model-releasesarxiv-cs-ai
7 Jul 2026
Tutorials

Revealing Hidden Model Behaviors with Task-Specific Self-Reports

DGX agent

arXiv:2607.03640v1 Announce Type: cross Abstract: Fine-tuning can give a language model a hidden behavior--it may give false answers under a narrow condition, or give harmful advice only when a prompt

tutorialsarxiv-cs-ai
7 Jul 2026
Applications

Signal or Noise? Understanding Generative Models for Real-World Sensor Time Series

DGX agent

arXiv:2607.04245v1 Announce Type: cross Abstract: Generative models have changed how machine learning represents complex data distributions, especially in language and vision, yet many real-world syst

applicationsarxiv-cs-ai
7 Jul 2026
Model Releases

Stacked LoRA for Subject-Adaptive EEG Foundation Models in Motor Imagery Decoding

DGX agent

arXiv:2607.03094v1 Announce Type: new Abstract: Electroencephalography (EEG) decoding for brain-computer interfaces (BCIs) faces a major challenge: substantial inter-subject variability limits effecti

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Training Hybrid Block Diffusion Language Models with Partial Bidirectionality

DGX agent

arXiv:2607.02805v1 Announce Type: cross Abstract: High-throughput long-context generation is one of the central challenges for large language models. Generation is typically memory-bandwidth-bound rat

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

UNIVERSE: Unified Video Action Models for Autonomous Driving with Flexible Mask-Modulated Modality Generation

DGX agent

arXiv:2607.05133v1 Announce Type: new Abstract: World Action Models (WAMs) have shown strong potential for improving action generalization in autonomous driving by using future video prediction as den

agentsarxiv-cs-cv
7 Jul 2026
Model Releases

VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents

DGX agent

arXiv:2510.11098v5 Announce Type: replace-cross Abstract: Recent advances in large audio language models (LALMs) have greatly enhanced multimodal conversational systems. However, existing benchmarks r

model-releasesarxiv-cs-cl
7 Jul 2026
Research

VISTA: Auditing Semantic Divergence in Vision-Language Models

DGX agent

arXiv:2607.02995v1 Announce Type: cross Abstract: Vision-language models can exhibit visual concept-conditioned divergence: given images containing demographic features, corporate logos, or ideologica

researcharxiv-cs-ai
7 Jul 2026
Applications

WAM4D: Fast 4D World Action Model via Spatial Register Tokens

DGX agent

arXiv:2606.14048v2 Announce Type: replace Abstract: World action models (WAMs) have recently shown promise in jointly modeling future observations and executable robot actions. However, most existing

applicationsarxiv-cs-cv
7 Jul 2026
Model Releases

Which Algorithm Specification Formats Help Language Models Implement Machine Learning Algorithms?

DGX agent

arXiv:2607.03158v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to implement algorithms from research manuscripts, but papers often leave implementation choices im

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Wrong Before Right: Late Rescue and Interface Failure in Aligned Language Models

DGX agent

arXiv:2607.04640v1 Announce Type: new Abstract: We study how correctness is assembled inside aligned language models, not only whether the final answer is right. Using layer-wise difference-in-differe

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

Claude Opus 4.8 and Sonnet 5 seem worse at tool calls than older models, likely due to post-training that assumes Claude Code-like harnesses as targets (Armin Ronacher/Armin Ronacher's Thoughts and Writings)

DGX agent

Armin Ronacher / Armin Ronacher's Thoughts and Writings: Claude Opus 4.8 and Sonnet 5 seem worse at tool calls than older models, likely due to post-training that assumes Claude Code-like harnesses as

model-releasestechmeme
6 Jul 2026
Model Releases

GLM-5.2 is now selectable in Claude Code via Hugging Face🤗 Inference Providers + hf-claude. Open models are becoming easier to plug directl…

DGX agent

GLM-5.2, an open-source model available through Hugging Face, can now be selected and used within Claude Code through Hugging Face Inference Providers and the hf-claude integration. This development d

model-releasesclem-delangue--x
3 Jul 2026
Research

Interpreting Global Perturbation Robustness of Image Models using Axiomatic Spectral Importance Decomposition

DGX agent

arXiv:2408.01139v4 Announce Type: replace Abstract: Perturbation robustness evaluates the vulnerabilities of models, arising from a variety of perturbations, such as data corruptions and adversarial a

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Meta to release new AI model with advanced coding capabilities ‘soon’

DGX agent

Meta Platforms Inc. is gearing up to release a new version of its flagship Muse Spark artificial intelligence model. Alexandr Wang, the company’s chief AI officer, wrote on X today that the update wil

model-releasessiliconangle
3 Jul 2026
Model Releases

MMBench-Live: A Continuously Evolving Benchmark for Multimodal Models

DGX agent

arXiv:2607.01813v1 Announce Type: cross Abstract: Evaluation benchmarks are essential for assessing vision-language models (VLMs), but most multimodal benchmarks are static, making them vulnerable to

model-releasesarxiv-cs-ai
3 Jul 2026
Research

OntoLearner: A Modular Python Library for Ontology Learning with Large Language Models

DGX agent

arXiv:2607.01977v1 Announce Type: new Abstract: Ontology learning (OL) aims to automatically construct structured knowledge models from text, yet progress remains fragmented across methods, domains, a

researcharxiv-cs-ai
3 Jul 2026
Agents

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down yo…

DGX agent

Prompt engineering is costing you money. Learn how to fine-tune your models. Stuffing a lot of text into every single API call slows down your app because the model has to process all those tokens bef

agentsfireworks-ai--x
3 Jul 2026
Model Releases

Psychological Imagination Networks Show Cross-Population Centrality and Clustering Alignment in Humans That Large Language Models Fail to Replicate

DGX agent

arXiv:2510.04391v5 Announce Type: replace Abstract: Mental imagery vividness is a stable individual trait, yet whether imagined scenarios share relational structure across human and synthetic large la

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns (The Information)

DGX agent

The Information: Sources: Alibaba has banned employees from using Claude Code and asked them to remove all Claude models from their work computers, citing security concerns — Alibaba Group has banned

model-releasestechmeme
3 Jul 2026
Model Releases

The Wiola Architecture for Efficient Small Language Models

DGX agent

arXiv:2607.01394v1 Announce Type: new Abstract: We present Wiola, a fully original Small Language Model (SLM) architecture built from first principles, sharing no structural lineage with any existing

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

DGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

model-releasesethan-mollick--x
3 Jul 2026
Model Releases

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few…

DGX agent

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few months is *Qwen 27b*. Our ML/AI engineering teams are have

model-releasesclem-delangue--x
3 Jul 2026
Research

3D Point World Models: Point Completion Enables More Accurate Dynamics Learning

DGX agent

arXiv:2607.00148v1 Announce Type: cross Abstract: Learning predictive models of the world enables robotic control through planning, potentially allowing robots to improvise solutions on new tasks. How

researcharxiv-cs-cv
2 Jul 2026
Model Releases

AD-MPCC: Adaptive Differentiable Model Predictive Contouring Control for Autonomous Racing

DGX agent

arXiv:2607.00141v1 Announce Type: new Abstract: This paper presents Adaptive Differentiable Model Predictive Contouring Control (AD-MPCC), a framework for autonomous racing that integrates differentia

model-releasesarxiv-cs-ro
2 Jul 2026
Research

Device Passport: Enabling Spatio-Temporal Pretrained Models to Generalize Across Input Layouts

DGX agent

arXiv:2607.00249v1 Announce Type: new Abstract: New device layouts pose a challenging modeling problem due to the lack of large datasets for each specific layout. Biosignal foundation models offer a p

researcharxiv-cs-lg
2 Jul 2026
Model Releases

DriveVA: Video Action Models are Zero-Shot Drivers

DGX agent

arXiv:2604.04198v2 Announce Type: replace Abstract: Generalization is a central challenge in autonomous driving, as real-world deployment requires robust performance under unseen scenarios, sensor dom

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

EmbodimentSemantic: A Spatial Scene-Graph Dataset and Benchmark for Vision-Language Models on Embodied Manipulation Trajectories

DGX agent

arXiv:2607.00020v1 Announce Type: new Abstract: Spatial grounding remains a key limitation of vision-language-action (VLA) systems for robotic manipulation. While current models can recognize objects

model-releasesarxiv-cs-ro
2 Jul 2026
Model Releases

Foundation Models vs. Radiomics for Lung Computed Tomography: A Benchmark of Feature Extractors, Classification Heads, and Segmentation Choices

DGX agent

arXiv:2607.01001v1 Announce Type: new Abstract: Radiomics is the established approach for CT-based lung cancer phenotyping, yet comparisons with foundation models rarely isolate contributions of featu

model-releasesarxiv-cs-cv
2 Jul 2026
← Previous
1…101102103104105…1261
Next →