AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
13 Apr 2026

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning

ResearchDGX agent

arXiv:2604.09529v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) achieve strong multimodal reasoning but frequently exhibit hallucinations and incorrect responses with high certa

11 Apr 2026

Great breakdown of how model providers are platformizing their AI/agents. A lot of people will take the convenience of going all in on a pro…

AgentsDGX agent

Great breakdown of how model providers are platformizing their AI/agents. A lot of people will take the convenience of going all in on a provider, but they will be locked in and giving up data control

Tesla friends, here’s some fun news: Tesla is doing a final Signature Series run of Plaid Model S and X. 250 S, 100 X (6-seat only). Invite-…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
IndustryDGX agent

Tesla friends, here’s some fun news: Tesla is doing a final Signature Series run of Plaid Model S and X. 250 S, 100 X (6-seat only). Invite-only so if you didn’t get the email you can’t buy one. Celeb

10 Apr 2026

A comparative analysis of machine learning models in SHAP analysis

TutorialsDGX agent

arXiv:2604.07258v1 Announce Type: new Abstract: In this growing age of data and technology, large black-box models are becoming the norm due to their ability to handle vast amounts of data and learn i

A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answering

Model ReleasesDGX agent

arXiv:2604.07274v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in medical question answering; however, purely parametric models often suffer from

Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06201v1 Announce Type: cross Abstract: While most reading comprehension benchmarks for LLMs focus on factual information that can be answered by localizing specific textual evidence, many r

CAFP: A Post-Processing Framework for Group Fairness via Counterfactual Model Averaging

SafetyDGX agent

arXiv:2604.07009v1 Announce Type: new Abstract: Ensuring fairness in machine learning predictions is a critical challenge, especially when models are deployed in sensitive domains such as credit scori

Deep Agents Deploy: an open alternative to Claude Managed Agents

Model ReleasesDGX agent

LangChain launched **Deep Agents Deploy** in beta as an open-source, model-agnostic alternative to Anthropic's Claude Managed Agents. It is designed to be the fastest way to deploy a model-agnosti...

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

SafetyDGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

Do We Need Distinct Representations for Every Speech Token? Unveiling and Exploiting Redundancy in Large Speech Language Models

ResearchDGX agent

arXiv:2604.06871v1 Announce Type: cross Abstract: Large Speech Language Models (LSLMs) typically operate at high token rates (tokens/s) to ensure acoustic fidelity, yet this results in sequence length

Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models

ResearchDGX agent

arXiv:2604.08456v1 Announce Type: cross Abstract: Despite rapid progress, pretrained vision-language models still struggle when answers depend on tiny visual details or on combining clues spread acros

Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization

SafetyDGX agent

arXiv:2604.08476v1 Announce Type: new Abstract: Multimodal reasoning models (MRMs) trained with reinforcement learning with verifiable rewards (RLVR) show improved accuracy on visual reasoning benchma

FedSpy-LLM: Towards Scalable and Generalizable Data Reconstruction Attacks from Gradients on LLMs

Model ReleasesDGX agent

arXiv:2604.06297v1 Announce Type: cross Abstract: Given the growing reliance on private data in training Large Language Models (LLMs), Federated Learning (FL) combined with Parameter-Efficient Fine-Tu

Guiding a Diffusion Model by Swapping Its Tokens

SafetyDGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathology

Model ReleasesDGX agent

arXiv:2604.08305v1 Announce Type: cross Abstract: Immunohistochemistry (IHC) is essential for assessing specific immune biomarkers like Human Epidermal growth-factor Receptor 2 (HER2) in breast cancer

I'm releasing the 34 slides on how we design and train best-in-class edge models at @liquidai I presented these slides yesterday at @aiDotEn…

ToolsDGX agent

I'm releasing the 34 slides on how we design and train best-in-class edge models at @liquidai I presented these slides yesterday at @aiDotEngineer They cover model architecture, pre-training, scaling

Latent Anomaly Knowledge Excavation: Unveiling Sparse Sensitive Neurons in Vision-Language Models

ResearchDGX agent

arXiv:2604.07802v1 Announce Type: new Abstract: Large-scale vision-language models (VLMs) exhibit remarkable zero-shot capabilities, yet the internal mechanisms driving their anomaly detection (AD) pe

Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classification

ResearchDGX agent

arXiv:2604.08333v1 Announce Type: new Abstract: The rise of multimodal large language models (MLLMs) has sparked an unprecedented wave of applications in the field of medical imaging analysis. However

LUMINA: Foundation Models for Topology Transferable ACOPF

SafetyDGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

Mind the Generative Details: Direct Localized Detail Preference Optimization for Video Diffusion Models

Local AiDGX agent

arXiv:2601.04068v3 Announce Type: replace Abstract: Aligning text-to-video diffusion models with human preferences is crucial for generating high-quality videos. Existing Direct Preference Otimization

Mitigating Entangled Steering in Large Vision-Language Models for Hallucination Reduction

ResearchDGX agent

arXiv:2604.07914v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable success across cross-modal tasks but remain hindered by hallucinations, producing textual

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

SafetyDGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

Self-Debias: Self-correcting for Debiasing Large Language Models

SafetyDGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

Model ReleasesDGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

The Human Condition as Reflected in Contemporary Large Language Models

ResearchDGX agent

arXiv:2604.06206v1 Announce Type: cross Abstract: This study seeks to uncover evidence of a latent structure in evolved human culture as it is refracted through contemporary large language models (LLM

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are signific…

TutorialsDGX agent

The pace at which useful things are shipping also seems to be accelerating. Model releases are coming faster, of course, but so are significant application and enterprise products (especially from Ant

VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents

Model ReleasesDGX agent

arXiv:2603.15118v2 Announce Type: replace Abstract: We introduce VAREX (VARied-schema EXtraction), a benchmark for evaluating multimodal foundation models on structured data extraction from government

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning

SafetyDGX agent

arXiv:2604.08168v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced robot manipulation through large-scale pretraining, but real-world deployment remains challenging due

When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoning

Model ReleasesDGX agent

arXiv:2604.08281v1 Announce Type: new Abstract: Large reasoning models (LRMs) have achieved strong performance enhancement through scaling test time computation, but due to the inherent limitations of

9 Apr 2026

First, Tesla canceled the Model 2—now it's working on a new small EV

IndustryDGX agent

After canceling its long-anticipated 'Model 2' affordable EV program in 2024 and pivoting toward robotaxis, Tesla is now reportedly developing an all-new compact electric SUV priced substantially b...

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) …

Model ReleasesDGX agent

the real future of the very best vertical products is Model/Harness Choice + Openness easy to deploy infra is nice (great release from Ant) but it’s not the lever that matters the most at all to build

8 Apr 2026

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new…

ResearchDGX agent

1/ today we're releasing muse spark, the first model from MSL. nine months ago we rebuilt our ai stack from scratch. new infrastructure, new architecture, new data pipelines. muse spark is the result

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, …

SafetyDGX agent

Common Failure Modes Break VLM-Powered OCR in Production. 🔁 Repetition Loops — model spirals into infinite whitespace, exhausts resources, cascades latency across your system 🛑 Recitation Errors — saf

15 Aug 2026

SF-based Vals, which develops evaluations and benchmarks to test AI models on real-world tasks, raised a 40M Series A led by a16z at a 400M valuation (Abhinaya Prabhu/Tech Funding News)

ApplicationsDGX agent

Abhinaya Prabhu / Tech Funding News: SF-based Vals, which develops evaluations and benchmarks to test AI models on real-world tasks, raised a 40M Series A led by a16z at a 400M valuation — - Vals AI r

14 Aug 2026

A hunch: Qwen3.8-27B's general knowledge got pruned (good, if true)

Model ReleasesDGX agent

I'm always testing an image prompt with a picture of a historic place in my hometown – a small but well known 250,000 people town in Germany. I'll just ask the model, in which City this photo has been

Adjustable Text-Guided Backdoor Attacks with Natural-Word Triggers on Multimodal Pretrained Models

ApplicationsDGX agent

arXiv:2604.05809v2 Announce Type: replace-cross Abstract: This paper presents Text-Guided Backdoor (TGB), an adjustable backdoor attack against multimodal pretrained models that uses natural-word trig

Alibaba releases weights for Qwen3.8 models under Apache 2.0 license, including Qwen3.8-27B, which it says beats Qwen3.7-Plus and excels in real-world coding (@alibaba_qwen)

ApplicationsDGX agent

@alibaba_qwen: Alibaba releases weights for Qwen3.8 models under Apache 2.0 license, including Qwen3.8-27B, which it says beats Qwen3.7-Plus and excels in real-world coding — We promised open weights

Automated Design Optimization via Strategic Search with Large Language Models

SafetyDGX agent

arXiv:2511.22651v2 Announce Type: replace-cross Abstract: Optimization methods have long advanced many fields, yet they struggle when faced with design problems where the search space and design param

Comment on 'Modeling rapid language learning by distilling Bayesian priors into artificial neural networks'

ResearchDGX agent

arXiv:2608.12974v1 Announce Type: cross Abstract: McCoy & Griffiths (2025, henceforth M&G) suggest that a Bayesian prior can be distilled into Artificial Neural Networks (ANNs) through Model-Agnostic

DomusFM: A Foundation Model for Event-Based Behavioral Monitoring in Smart-Homes

ApplicationsDGX agent

arXiv:2602.01910v2 Announce Type: replace Abstract: Smart-home sensor-based behavioral monitoring holds significant potential for healthcare, independent living, and early detection of functional or c

Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese

Model ReleasesDGX agent

arXiv:2608.12373v1 Announce Type: new Abstract: Large language models are increasingly used in strategic and advisory contexts, yet their safety alignment is typically evaluated in English only. We te

FlashDrive: Flash Vision-Language-Action Inference for Autonomous Driving

Model ReleasesDGX agent

arXiv:2608.12932v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models promise to bring end-to-end reasoning to autonomous driving, but their computational cost remains far too high for r

Foundation models for movement data: Are they ready for prime-time?

ResearchDGX agent

arXiv:2608.13316v1 Announce Type: cross Abstract: Foundation models (FMs) trained on large-scale accelerometer data have been proposed as general-purpose feature extractors for health monitoring, but

GEM: A Generative Embedding Model Bridging Reasoning and Retrieval

ResearchDGX agent

arXiv:2608.13200v1 Announce Type: cross Abstract: Modern LLMs excel at reasoning and instruction following, enabling users to express complex and diverse information needs. However, conventional retri

H3 as a single-image edit model

TutorialsDGX agent

Minimax H3 can be used as an image-editing model if we generate a single frame. Here are some collages based on AI-generated references with workflows. Each edit takes about 8 secs on a RTX 5090. The

I’ve tried Cursor, Claude Code, Ollama, etc. — but I still don’t know how to use AI effectively for coding

Model ReleasesDGX agent

I’ve experimented with Cursor, Antigravity, Claude Code/CLI, Ollama, Gemma 4B, MiniMax, Hermes Agent, and different local/cloud models. My problem isn’t knowing what these tools are—I don't know how t

MASCOT: Model-Aware Submodular Coverage for Composite-Attribute Text-to-Image Retrieval

ResearchDGX agent

arXiv:2608.12532v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are highly effective in retrieving semantically relevant images. However, in practice, relevance alone is often insuffic

Numeracy in Large Language Models: Fundamental Limitations and Paths to Improvement

ResearchDGX agent

arXiv:2608.13129v1 Announce Type: new Abstract: Large language models (LLMs) achieve strong results on mathematical reasoning benchmarks yet remain unreliable on elementary numerical tasks, including

Perturbation-based Regional Interpretability through Subtraction Mapping (PRISM): naming-error dissociations in language models and post-stroke aphasia

ResearchDGX agent

arXiv:2608.12717v1 Announce Type: cross Abstract: Mechanistic interpretability of large language models lacks spatially resolved, falsifiable tools for testing whether internal components are speciali

REHEARSE: Experiential Rehearsal for Verbal Confidence Calibration in Large Language Models

SafetyDGX agent

arXiv:2508.14390v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often express verbal confidence that is poorly aligned with actual correctness, limiting their reliability in saf

The Embedder's Dilemma: LLMs Are Better, but at What Cost?

Model ReleasesDGX agent

arXiv:2608.12875v1 Announce Type: new Abstract: Should you replace your text-embedding pipeline with a large language model? We answer this with a controlled, cost-aware comparison of ten LLMs across

UniTexture: Cross-Task Universal Adversarial Textures for Vision-Language-Action Models

SafetyDGX agent

arXiv:2608.13453v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as generalist robotic policies capable of following diverse language instructions and performing a wi

Unlearning at Scale: State-Exact Trace-Preserving Deletion in Billion-Parameter Language Models

Model ReleasesDGX agent

arXiv:2508.12220v2 Announce Type: replace-cross Abstract: Can a prospectively instrumented training continuation reproduce a deletion counterfactual exactly after selected examples leave its replay da

When Explanations Betray Backdoors: Black-Box Auditing for Language Model Classifiers

ResearchDGX agent

arXiv:2608.12623v1 Announce Type: new Abstract: Language model classifiers with explanations are used for moderation, routing, topic triage, and low-resource annotation. We study black-box auditing wh

Z.ai debuts GLM-5.3 with long-horizon coding, cybersecurity upgrades

Model ReleasesDGX agent

Chinese artificial intelligence developer Z.ai Co. today debuted GLM-5.3, an open-source large language model that set records across several popular benchmarks. The LLM is based on an algorithm calle

13 Aug 2026

A Theoretical Framework for Modular Learning of Robust Generative Models

ApplicationsDGX agent

arXiv:2602.17554v3 Announce Type: replace Abstract: Training large-scale generative models is resource-intensive and relies heavily on heuristic dataset weighting. We address two fundamental questions

COLORA: Efficient Fine-Tuning for Convolutional Models with a Study Case on Optical Coherence Tomography Image Classification

Model ReleasesDGX agent

arXiv:2505.18315v3 Announce Type: replace-cross Abstract: We introduce extbf{CoLoRA} (Convolutional Low-Rank Adaptation), a parameter-efficient fine-tuning method for convolutional neural networks (CN

Deep Activity Model: A Generative Approach for Human Mobility Pattern Synthesis

Local AiDGX agent

arXiv:2405.17468v3 Announce Type: replace-cross Abstract: Human mobility plays a crucial role in transportation, urban planning, and public health, but current approaches face important limitations. E

Dual-Model Sentiment Analysis of Consumer Reviews in the Retail Coffee Sector Using Machine Learning and Deep Learning Approaches

ApplicationsDGX agent

arXiv:2608.12007v1 Announce Type: cross Abstract: Consumer reviews play an important role in shaping brand perception and business strategies, particularly in service-driven industries such as retail

Foresight Without Seeing: Latent Futures for World Action Models

SafetyDGX agent

arXiv:2608.11605v1 Announce Type: new Abstract: World Action Models (WAMs) couple future visual prediction with robot action generation, enabling policies to model how the physical world evolves durin

← Previous
1…122123124125126…1009
Next →