AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,515 results
Model Releases

CT-DeltaBench: A Benchmark for Longitudinal 3D Medical Imaging Difference Reporting with Vision-Language Models

DGX agent

arXiv:2608.11534v1 Announce Type: new Abstract: In medical imaging, the clinical value of Computed Tomography (CT) lies not only in depicting current disease status, but crucially in enabling longitud

model-releasesarxiv-cs-cl
13 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?

DGX agent

arXiv:2603.00056v2 Announce Type: replace-cross Abstract: STEM Mental models can play a critical role in assessing students' conceptual understanding of a topic. They not only offer insights into what

researcharxiv-cs-ai
13 Aug 2026
Model Releases

Orientation, not magnitude: the causal structure of task-vector interference in merged language models

DGX agent

arXiv:2608.11797v1 Announce Type: new Abstract: Model merging by task arithmetic works until it doesn't, and the field diagnoses why with magnitudes: layerwise representation bias, deviations from cro

model-releasesarxiv-cs-lg
13 Aug 2026
Model Releases

Poor Man's Agentic Modeling: Simulating Large LLM-Agent Societies on a Laptop

DGX agent

arXiv:2608.11215v1 Announce Type: new Abstract: Simulating societies of many large language model (LLM) agents is expensive, yet the questions asked of such simulations are usually macroscopic: phase

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Qwen-MusicAVQA-7B: A Multimodal Model for Music Audio-Visual QA

DGX agent

arXiv:2608.11329v1 Announce Type: cross Abstract: A common approach to adding audio to a vision-language model is to train or adapt a large omni-modal system. We show that a lightweight alternative ca

model-releasesarxiv-cs-cv
13 Aug 2026
Model Releases

Semantic Lenia: Emergence of Homeostatic Solitons within the Semantic Space of Large Language Models

DGX agent

arXiv:2608.11657v1 Announce Type: cross Abstract: We introduce Semantic Lenia, an artificial life framework that transforms Large Language Model (LLM) inference from a static optimization problem into

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Writer launches major agentic AI improvements with Palmyra X6 flagship model

DGX agent

Generative artificial intelligence startup Writer Inc. today announced the release of its next-generation flagship model, Palmyra X6, designed to deliver frontier-level performance for marketing and r

model-releasessiliconangle
13 Aug 2026
Local Ai

Certify or Refuse: A Cross-Model Map for Selective Risk Control with Coverage Floors under Covariate Shift

DGX agent

arXiv:2608.10893v1 Announce Type: new Abstract: Certified selective predictors attain whatever coverage they attain; operators impose an automation floor: answer at least a eta-fraction of shifted tar

local-aiarxiv-cs-cl
12 Aug 2026
Safety

Hidden in Plain Sight: Diffusion-Based Unrestricted Robotic Attacks on Vision-Language-Action Models

DGX agent

arXiv:2608.10393v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong capabilities in controlling robots across diverse manipulation tasks. However, their adversarial r

safetyarxiv-cs-ai
12 Aug 2026
Research

HQ-DM: Single Hadamard Transformation-Based Quantization-Aware Training for Low-Bit Diffusion Models

DGX agent

arXiv:2512.05746v3 Announce Type: replace Abstract: Diffusion models have demonstrated significant applications in the field of image generation. However, their high computational and memory costs pos

researcharxiv-cs-cv
12 Aug 2026
Model Releases

Locally Deployable Small Language Models for Emergency Department Decision Support: A Systematic Benchmark of Fine-Tuning Strategies

DGX agent

arXiv:2608.10273v1 Announce Type: cross Abstract: Deploying large language models (LLMs) for decision support in emergency departments (EDs) faces two major challenges: privacy risks of transmitting p

model-releasesarxiv-cs-ai
12 Aug 2026
Industry

Palo Alto Networks to run OpenAI cyber models inside customer networks

DGX agent

Palo Alto Networks Inc. said today its Unit 42 consulting arm will put OpenAI Group PBC’s frontier cyber models to work inside customer environments, expanding a service it launched earlier this year

industrysiliconangle
12 Aug 2026
Tutorials

ReCBM: Uncertainty-Gated Relational Reasoning for Concept Bottleneck Models

DGX agent

arXiv:2608.10004v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide an interpretable framework by grounding predictions in human-understandable concepts, enabling semantic inspect

tutorialsarxiv-cs-ai
12 Aug 2026
Safety

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

DGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

safetyarxiv-cs-ai
12 Aug 2026
Agents

Activation Probes Surface Code-Security Signals that the Model's Output Misses

DGX agent

arXiv:2608.09643v1 Announce Type: cross Abstract: AI coding agents now write a growing share of production code, and human security review does not scale at the rate code is generated. The agents in w

agentsarxiv-cs-lg
11 Aug 2026
Model Releases

An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography

DGX agent

arXiv:2608.07651v1 Announce Type: new Abstract: Large language models (LLMs) show promise in medical image interpretation but suffer from hallucination, limited accuracy, and run-to-run inconsistency.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CosmosAlign: Adapting a World Foundation Model for Generative Traffic Video Forecasting

DGX agent

arXiv:2608.07693v1 Announce Type: cross Abstract: Generative traffic video forecasting aims to synthesize long-horizon, temporally coherent future videos of traffic scenes from a short observation his

model-releasesarxiv-cs-ai
11 Aug 2026
Research

DialectS2S: End-to-End Speech Dialogue Modeling for Low-Resource Chinese Dialects

DGX agent

arXiv:2608.08067v1 Announce Type: cross Abstract: Current end-to-end speech dialogue models are primarily optimized for mainstream languages and remain limited in low-resource dialect scenarios due to

researcharxiv-cs-ai
11 Aug 2026
Tutorials

Ensemble learning of pathology foundation models for precision oncology

DGX agent

arXiv:2508.16085v2 Announce Type: replace Abstract: Histopathology is essential for cancer diagnosis and treatment selection, and pathology foundation models learn visual representations from whole-sl

tutorialsarxiv-cs-cv
11 Aug 2026
Model Releases

FemWear: A Specialized Wearable Foundation Model for Women's Health

DGX agent

arXiv:2608.08244v1 Announce Type: new Abstract: General wearable foundation models are pretrained across broad sensor streams and populations, but are not designed around women's-health tasks. We intr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

FitAQA: A Benchmark of Fitness Action Quality Assessment for Multimodal Large Language Models

DGX agent

arXiv:2608.08736v1 Announce Type: new Abstract: Fitness Action Quality Assessment (AQA) is important for intelligent sports training, yet the capabilities of Multimodal Large Language Models (MLLMs) i

model-releasesarxiv-cs-ai
11 Aug 2026
Research

From Independent to Correlated Diffusion: Generalized Generative Modeling with Probabilistic Computers

DGX agent

arXiv:2603.27996v2 Announce Type: replace Abstract: Diffusion models have emerged as a powerful framework for generative tasks in deep learning. They decompose generative modeling into two computation

researcharxiv-cs-lg
11 Aug 2026
Model Releases

Fusion Training for Mathematical Generalization in Large Language Models

DGX agent

arXiv:2608.09893v1 Announce Type: cross Abstract: Thinking Mode Fusion (TMF) enables large language models to support both concise responses and long-form reasoning by unifying a non-thinking mode and

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

GWM-VLA: Geometry-Aware Latent World Modeling for Vision-Language-Action Learning

DGX agent

arXiv:2608.07619v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong robotic manipulation performance but often degrade under visual and environmental shifts. Latent worl

local-aiarxiv-cs-ro
11 Aug 2026
Model Releases

Local Benchmark : Muse Glimmer 30B vs Qwen 3.6 27B vs Gemma4 31B (and many other models and finetunes)

DGX agent

Needs a lot of requests compared to Qwen (almost twice) and Gemma (almost x3). Final score is fine, even though it is 'not a coding model' https://wonderrico.github.io/local_llm_benchmark/benchmark-ma

model-releasesr-localllama
11 Aug 2026
Model Releases

☁️Mistral is bringing together the inference infrastructure, open models, and long-term commitments Europe needs to control its AI future, a…

DGX agent

☁️Mistral is bringing together the inference infrastructure, open models, and long-term commitments Europe needs to control its AI future, and setting a roadmap for the world. 🧵: https://mistral.ai/ne

model-releasesmistral-ai--x
11 Aug 2026
Research

MoE-Prism: Disentangling Monolithic Experts for Elastic MoE Services via Model-System Co-Designs

DGX agent

arXiv:2510.19366v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) scales model capacity through sparse activation, and is becoming an important architecture for large language models (LLMs)

researcharxiv-cs-cl
11 Aug 2026
Model Releases

MonitorBench: A Comprehensive Benchmark for Chain-of-Thought Monitorability in Large Language Models

DGX agent

arXiv:2603.28590v3 Announce Type: replace Abstract: Large language models (LLMs) can generate chains of thought (CoTs) that are not always causally responsible for their final outputs. When such a mis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Omni2LoRA: Coherence-Preserving Parametric Memory for Efficient Omni Language Models

DGX agent

arXiv:2608.09227v1 Announce Type: new Abstract: Omnimodal language models (OLMs) enable unified audio-visual understanding, but processing long joint token sequences makes inference computationally pr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

PragMatch: Separating Pragmatic Incongruity from Cross-Modal Mismatch in Large Vision-Language Models

DGX agent

arXiv:2608.09772v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated strong performance on multimodal benchmarks, yet it remains unclear whether they genuinely reason

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Scaling Audio Models Efficiently: A Joint Study of Compute Constraints and Optimization Behavior

DGX agent

arXiv:2606.22790v2 Announce Type: replace-cross Abstract: In this paper, we investigate the tradeoffs between compute allocation and model performance for two speech processing tasks: Automatic Speech

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

DGX agent

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries,

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

The Evolution of Mixture-of-Experts Architectures in Large Language Models: Routing, Topology, Load Balancing, and Expert Parallelism

DGX agent

arXiv:2608.08650v1 Announce Type: new Abstract: Mixture-of-Experts models increase parameter capacity while keeping the computation activated by each token bounded, but their architectural evolution c

model-releasesarxiv-cs-cl
11 Aug 2026
Research

TokenPrint: A Calibrated Token-Space Fingerprint for Language-Model Provenance

DGX agent

arXiv:2608.08139v1 Announce Type: new Abstract: Establishing the provenance of a language model---including its base checkpoint and possible overlap in training distributions---is a governance challen

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Research

World Simulator: Queer Erotica and the Absurdity of AI Video Models That Promise the World

DGX agent

arXiv:2608.07510v1 Announce Type: cross Abstract: Increasingly, AI video models are marketed as 'world simulators,' suggesting their ability to model infinite realities. Despite such claims, these mod

researcharxiv-cs-cv
11 Aug 2026
Research

ZOMP: Zeroth-Order Multi-Modal Prompt Tuning for Vision-Language Models

DGX agent

arXiv:2608.08060v1 Announce Type: new Abstract: Fine-tuning vision-language models such as CLIP typically requires backpropagation (BP) through the full model, which is infeasible when only forward-pa

researcharxiv-cs-cv
11 Aug 2026
Model Releases

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

DGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

ArchEGraph: A Large-Scale Graph Dataset for Geometry-Topology-Physics Aligned Building Energy Modeling

DGX agent

arXiv:2608.06772v1 Announce Type: new Abstract: Accurate estimation of building energy use is essential for achieving carbon neutral and sustainable buildings. To better understand the influence of de

model-releasesarxiv-cs-lg
10 Aug 2026
Safety

Do 3D Medical Foundation Models See Through MRI Artifacts? A Controlled Study of Representation Robustness

DGX agent

arXiv:2608.06613v1 Announce Type: cross Abstract: Self-supervised 3D medical foundation models are increasingly used as general-purpose feature extractors, yet their sensitivity to MRI artifacts remai

safetyarxiv-cs-ai
10 Aug 2026
Research

Flow-Corrected Shape Optimization: Taming Manifold Drift in High-Dimensional 3D Models

DGX agent

arXiv:2608.07199v1 Announce Type: new Abstract: Optimizing 3D shapes within the latent spaces of deep generative models is fundamental to computer assisted engineering, yet remains prone to a critical

researcharxiv-cs-cv
10 Aug 2026
Research

Gaze Behavior in Visual World Experiments Can be Modeled With Off-the-shelf Language-Vision Encoders

DGX agent

arXiv:2608.07282v1 Announce Type: new Abstract: The recent advances in neural language models have also spurred much work in computational psycholinguistics, asking whether neural LMs are also promisi

researcharxiv-cs-cl
10 Aug 2026
Safety

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

DGX agent

arXiv:2608.06799v1 Announce Type: cross Abstract: Learning structured and control-relevant latent representations remains a key challenge for world models. Recent JEPA-based world models learn action-

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Multi Codec Discrete Diffusion Model for Text Guided Speech Inpainting and Editing

DGX agent

arXiv:2608.06424v1 Announce Type: cross Abstract: Speech recordings often contain missing, corrupted, or incorrect regions that must be reconstructed or modified without re-synthesizing the entire utt

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

PHASE-Tree: Modeling Character-State Evolution in Long-Horizon Role-Playing Dialogue

DGX agent

arXiv:2608.06975v1 Announce Type: cross Abstract: Long-horizon role-playing demands that characters remain recognizable as they evolve with the narrative. Yet existing work falls short on two fronts:

model-releasesarxiv-cs-ai
10 Aug 2026
Tutorials

UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling

DGX agent

arXiv:2608.07409v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) have emerged as a principled framework for self-supervised learning of world models in compact latent s

tutorialsarxiv-cs-cv
10 Aug 2026
Model Releases

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CE…

DGX agent

Very cool idea to have agents design complex systems by searching over the model structure itself. New research from Sakana AI introduces CEDAR, which uses LLM agents to write, simulate, and refine sy

model-releasesdair-ai--x
10 Aug 2026
← Previous
1…96979899100…1261
Next →