AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
Industry

Tesla friends, here’s some fun news: Tesla is doing a final Signature Series run of Plaid Model S and X. 250 S, 100 X (6-seat only). Invite-…

DGX agent

Tesla friends, here’s some fun news: Tesla is doing a final Signature Series run of Plaid Model S and X. 250 S, 100 X (6-seat only). Invite-only so if you didn’t get the email you can’t buy one. Celeb

industryelon-musk--x
11 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

A comparative analysis of machine learning models in SHAP analysis

DGX agent

arXiv:2604.07258v1 Announce Type: new Abstract: In this growing age of data and technology, large black-box models are becoming the norm due to their ability to handle vast amounts of data and learn i

tutorialsarxiv-cs-lg
10 Apr 2026
Model Releases

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

DGX agent

arXiv:2604.07274v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models

DGX agent

arXiv:2604.06201v1 Announce Type: cross Abstract: While most reading comprehension benchmarks for LLMs focus on factual information that can be answered by localizing specific textual evidence, many r

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

DGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Deep Agents Deploy: an open alternative to Claude Managed Agents

DGX agent

LangChain launched **Deep Agents Deploy** in beta as an open-source, model-agnostic alternative to Anthropic's Claude Managed Agents. It is designed to be the fastest way to deploy a model-agnosti...

model-releasesharrison-chase--x
10 Apr 2026
Safety

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

DGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

safetyarxiv-cs-cl
10 Apr 2026
Research

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

DGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

researcharxiv-cs-ai
10 Apr 2026
Research

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

DGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

researcharxiv-cs-cl
10 Apr 2026
Safety

Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization

DGX agent

arXiv:2604.08476v1 Announce Type: new Abstract: Multimodal reasoning models (MRMs) trained with reinforcement learning with verifiable rewards (RLVR) show improved accuracy on visual reasoning benchma

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

FedSpy-LLM: Towards Scalable and Generalizable Data Reconstruction Attacks from Gradients on LLMs

DGX agent

arXiv:2604.06297v1 Announce Type: cross Abstract: Given the growing reliance on private data in training Large Language Models (LLMs), Federated Learning (FL) combined with Parameter-Efficient Fine-Tu

model-releasesarxiv-cs-lg
10 Apr 2026
Safety

Guiding a Diffusion Model by Swapping Its Tokens

DGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathology

DGX agent

arXiv:2604.08305v1 Announce Type: cross Abstract: Immunohistochemistry (IHC) is essential for assessing specific immune biomarkers like Human Epidermal growth-factor Receptor 2 (HER2) in breast cancer

model-releasesarxiv-cs-cv
10 Apr 2026
Tools

I'm releasing the 34 slides on how we design and train best-in-class edge models at @liquidai I presented these slides yesterday at @aiDotEn…

DGX agent

I'm releasing the 34 slides on how we design and train best-in-class edge models at @liquidai I presented these slides yesterday at @aiDotEngineer They cover model architecture, pre-training, scaling

toolsswyx--x
10 Apr 2026
Research

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2604.07802v1 Announce Type: new Abstract: Large-scale vision-language models (VLMs) exhibit remarkable zero-shot capabilities, yet the internal mechanisms driving their anomaly detection (AD) pe

researcharxiv-cs-cv
10 Apr 2026
Research

Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification

DGX agent

arXiv:2604.08333v1 Announce Type: new Abstract: The rise of multimodal large language models (MLLMs) has sparked an unprecedented wave of applications in the field of medical imaging analysis. However

researcharxiv-cs-cv
10 Apr 2026
Safety

LUMINA: Foundation Models for Topology Transferable ACOPF

DGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

safetyarxiv-cs-lg
10 Apr 2026
Local Ai

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models

DGX agent

arXiv:2601.04068v3 Announce Type: replace Abstract: Aligning text-to-video diffusion models with human preferences is crucial for generating high-quality videos. Existing Direct Preference Otimization

local-aiarxiv-cs-cv
10 Apr 2026
Research

Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction

DGX agent

arXiv:2604.07914v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable success across cross-modal tasks but remain hindered by hallucinations, producing textual

researcharxiv-cs-cv
10 Apr 2026
Safety

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

DGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

safetyarxiv-cs-cl
10 Apr 2026
Safety

Self-Debias: Self-correcting for Debiasing Large Language Models

DGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

DGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

model-releasesarxiv-cs-ai
10 Apr 2026
Research

The Human Condition as Reflected in Contemporary Large Language Models

DGX agent

arXiv:2604.06206v1 Announce Type: cross Abstract: This study seeks to uncover evidence of a latent structure in evolved human culture as it is refracted through contemporary large language models (LLM

researcharxiv-cs-ai
10 Apr 2026
Tutorials

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are signific…

DGX agent

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are significant application and enterprise products (especially from Ant

tutorialsethan-mollick--x
10 Apr 2026
Model Releases

VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents

DGX agent

arXiv:2603.15118v2 Announce Type: replace Abstract: We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

DGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning

DGX agent

arXiv:2604.08281v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of

model-releasesarxiv-cs-cl
10 Apr 2026
Industry

First, Tesla canceled the Model 2—now it's working on a new small EV

DGX agent

After canceling its long-anticipated 'Model 2' affordable EV program in 2024 and pivoting toward robotaxis, Tesla is now reportedly developing an all-new compact electric SUV priced substantially b...

industryars-technica
9 Apr 2026
Model Releases

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) …

DGX agent

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) but it’s not the lever that matters the most at all to build

model-releasesharrison-chase--x
9 Apr 2026
Research

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new…

DGX agent

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new architecture, new data pipelines. muse spark is the result

researchyann-lecun--x
8 Apr 2026
Safety

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, …

DGX agent

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, cascades latency across your system 🛑 Recitation Errors — saf

safetyjerry-liu--x
8 Apr 2026
Hardware

AI is becoming a new computing model for media and entertainment.

DGX agent

AI is becoming a new computing model for media and entertainment. AI is becoming a new computing model for media and entertainment. At the 2026 Runway AI Summit, NVIDIA’s Richard Kerris explored how r

hardwarecristobal-valenzuela--x
15 Aug 2026
Applications

SF-based Vals, which develops evaluations and benchmarks to test AI models on real-world tasks, raised a 40M Series A led by a16z at a 400M valuation (Abhinaya Prabhu/Tech Funding News)

DGX agent

Abhinaya Prabhu / Tech Funding News: SF-based Vals, which develops evaluations and benchmarks to test AI models on real-world tasks, raised a 40M Series A led by a16z at a 400M valuation — - Vals AI r

applicationstechmeme
15 Aug 2026
Model Releases

A hunch: Qwen3.8-27B's general knowledge got pruned (good, if true)

DGX agent

I'm always testing an image prompt with a picture of a historic place in my hometown – a small but well known 250,000 people town in Germany. I'll just ask the model, in which City this photo has been

model-releasesr-localllama
14 Aug 2026
Applications

Adjustable Text-Guided Backdoor Attacks with Natural-Word Triggers on Multimodal Pretrained Models

DGX agent

arXiv:2604.05809v2 Announce Type: replace-cross Abstract: This paper presents Text-Guided Backdoor (TGB), an adjustable backdoor attack against multimodal pretrained models that uses natural-word trig

applicationsarxiv-cs-lg
14 Aug 2026
Applications

Alibaba releases weights for Qwen3.8 models under Apache 2.0 license, including Qwen3.8-27B, which it says beats Qwen3.7-Plus and excels in real-world coding (@alibaba_qwen)

DGX agent

@alibaba_qwen: Alibaba releases weights for Qwen3.8 models under Apache 2.0 license, including Qwen3.8-27B, which it says beats Qwen3.7-Plus and excels in real-world coding — We promised open weights

applicationstechmeme
14 Aug 2026
Safety

Automated Design Optimization via Strategic Search with Large Language Models

DGX agent

arXiv:2511.22651v2 Announce Type: replace-cross Abstract: Optimization methods have long advanced many fields, yet they struggle when faced with design problems where the search space and design param

safetyarxiv-cs-ai
14 Aug 2026
Research

Comment on 'Modeling rapid language learning by distilling Bayesian priors into artificial neural networks'

DGX agent

arXiv:2608.12974v1 Announce Type: cross Abstract: McCoy & Griffiths (2025, henceforth M&G) suggest that a Bayesian prior can be distilled into Artificial Neural Networks (ANNs) through Model-Agnostic

researcharxiv-cs-cl
14 Aug 2026
Applications

DomusFM: A Foundation Model for Event-Based Behavioral Monitoring in Smart-Homes

DGX agent

arXiv:2602.01910v2 Announce Type: replace Abstract: Smart-home sensor-based behavioral monitoring holds significant potential for healthcare, independent living, and early detection of functional or c

applicationsarxiv-cs-ai
14 Aug 2026
Model Releases

Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese

DGX agent

arXiv:2608.12373v1 Announce Type: new Abstract: Large language models are increasingly used in strategic and advisory contexts, yet their safety alignment is typically evaluated in English only. We te

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

FlashDrive: Flash Vision-Language-Action Inference for Autonomous Driving

DGX agent

arXiv:2608.12932v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models promise to bring end-to-end reasoning to autonomous driving, but their computational cost remains far too high for r

model-releasesarxiv-cs-ai
14 Aug 2026
Research

Foundation models for movement data: Are they ready for prime-time?

DGX agent

arXiv:2608.13316v1 Announce Type: cross Abstract: Foundation models (FMs) trained on large-scale accelerometer data have been proposed as general-purpose feature extractors for health monitoring, but

researcharxiv-cs-lg
14 Aug 2026
Research

GEM: A Generative Embedding Model Bridging Reasoning and Retrieval

DGX agent

arXiv:2608.13200v1 Announce Type: cross Abstract: Modern LLMs excel at reasoning and instruction following, enabling users to express complex and diverse information needs. However, conventional retri

researcharxiv-cs-ai
14 Aug 2026
Tutorials

H3 as a single-image edit model

DGX agent

Minimax H3 can be used as an image-editing model if we generate a single frame. Here are some collages based on AI-generated references with workflows. Each edit takes about 8 secs on a RTX 5090. The

tutorialsr-stablediffusion
14 Aug 2026
Model Releases

I’ve tried Cursor, Claude Code, Ollama, etc. — but I still don’t know how to use AI effectively for coding

DGX agent

I’ve experimented with Cursor, Antigravity, Claude Code/CLI, Ollama, Gemma 4B, MiniMax, Hermes Agent, and different local/cloud models. My problem isn’t knowing what these tools are—I don't know how t

model-releasesr-ollama
14 Aug 2026
Research

MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval

DGX agent

arXiv:2608.12532v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are highly effective in retrieving semantically relevant images. However, in practice, relevance alone is often insuffic

researcharxiv-cs-cv
14 Aug 2026
Research

Numeracy in Large Language Models: Fundamental Limitations and Paths to Improvement

DGX agent

arXiv:2608.13129v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong results on mathematical reasoning benchmarks yet remain unreliable on elementary numerical tasks, including

researcharxiv-cs-ai
14 Aug 2026
Research

Perturbation-based Regional Interpretability through Subtraction Mapping (PRISM): naming-error dissociations in language models and post-stroke aphasia

DGX agent

arXiv:2608.12717v1 Announce Type: cross Abstract: Mechanistic interpretability of large language models lacks spatially resolved, falsifiable tools for testing whether internal components are speciali

researcharxiv-cs-cl
14 Aug 2026
← Previous
1…153154155156157…1262
Next →