AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
16 Aug 2026

Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)

Model ReleasesDGX agent

Including the rationalisation for the data below - this is a more robust version of an earlier post I did similar to this - explaining below: How I chose the comparisons The basic question I’m trying

Ollama works locally, but my coding-agent integration does not

Model ReleasesDGX agent

I’m running Ollama on an M4 Pro Mac with 48 GB unified memory and testing local coding models for repository work. Direct Ollama inference is fine. Qwen3.8 27B-MLX runs well enough for me, and the loc

14 Aug 2026

Beyond Correctness: Benchmarking and Aligning Response Behaviors in Hybrid-Thinking MLLMs

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2608.12781v1 Announce Type: new Abstract: Hybrid-thinking multimodal large language models (MLLMs) allow a single model to alternate between deliberative thinking and latency-efficient non-think

Do LLMs Beat Nash? Testing Decentralized Coordination in Self-Play Multi-Agent Games

Model ReleasesDGX agent

arXiv:2608.12547v1 Announce Type: cross Abstract: Large language model agents deployed without a central controller are often assumed to require communication to coordinate their actions. We ask what

Edit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZ

Model ReleasesDGX agent

arXiv:2608.13441v1 Announce Type: new Abstract: Although multimodal large language models (MLLMs) have shown substantial potential in visual understanding and graphic code generation, editing scientif

Foam-Agent: A Large Language Model-Based Multi-Agent Framework for Automating Computational Fluid Dynamics Workflows

AgentsDGX agent

arXiv:2505.04997v3 Announce Type: replace Abstract: Computational fluid dynamics (CFD) has been the main workhorse of computational physics, yet its steep learning curve and fragmented, multi-stage wo

From Atomic Evidence to Logical Composition: Structured Compositional Reasoning over Compound Answer Options

Model ReleasesDGX agent

arXiv:2608.12836v1 Announce Type: cross Abstract: Large language models often fail when answer options require combining atomic judgments under explicit logical operators, even when they judge the ind

How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures

Model ReleasesDGX agent

arXiv:2608.13267v1 Announce Type: cross Abstract: Existing vision-language model (VLM) benchmarks emphasize perception and reasoning accuracy (how well VLMs describe and reason about what they see in

LoRA-Diffusion: Parameter-Efficient Fine-Tuning via Low-Rank Trajectory Decomposition

Model ReleasesDGX agent

arXiv:2608.12328v1 Announce Type: new Abstract: Parameter-efficient fine-tuning methods such as LoRA have transformed the adaptation of large autoregressive language models, enabling task-specific cus

Poll results 6 months later: When will we have Opus level with 30b model? Optimists win!

Local AiDGX agent

Poll at beginning of the year: https://www.reddit.com/r/LocalLLaMA/comments/1qj935h/poll_when_will_we_have_a_30b_open_weight_model_as/ The least voted option, 6 months, wins in my opinion, with 18 vot

Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs

Model ReleasesDGX agent

arXiv:2608.12675v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely used to improve the performance of Large Language Models (LLMs) in answering user queries. Existing priva

QuISE: Defense against Typographic Attacks on VLMs via Query-Irrelevant Semantic Editing

ResearchDGX agent

arXiv:2608.13119v1 Announce Type: new Abstract: Typographic attacks pose a critical threat to vision-language models (VLMs) by injecting misleading text into images and causing models to rely on adver

QuoteBench: How Matched Scores Can Hide Command-Path Failures

Model ReleasesDGX agent

arXiv:2608.13547v1 Announce Type: new Abstract: LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot disti

Towards Socially Compliant Navigation in Deep Reinforcement Learning via Proxemics-Based Reward Modeling

ApplicationsDGX agent

arXiv:2608.12917v1 Announce Type: new Abstract: Developing effective robot navigation methods in crowded environments is essential for real-world applications. Although recent deep reinforcement learn

Virtual Temperature Sensors in Power Transformers Using Neural Ordinary Differential Equations

Model ReleasesDGX agent

arXiv:2608.13260v1 Announce Type: new Abstract: Accurate modeling and forecasting of power transformer thermal behavior are critical for reliability, asset lifetime, and optimized power system operati

Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance

Model ReleasesDGX agent

arXiv:2608.12323v1 Announce Type: cross Abstract: Specifying a penalty can paradoxically convert a legal obligation into a cost-benefit calculation that favors violation. We demonstrate that this enfo

13 Aug 2026

From Self-Normal-Positioning to Omni-Directional Tracking: Real-Time Surface Modeling Enabled Probe Tilt Control for Robotic Ultrasound Imaging

Local AiDGX agent

arXiv:2608.11409v1 Announce Type: new Abstract: Ultrasound (US) provides real-time, radiation-free imaging, but the image quality depends strongly on how the probe is oriented against the patient body

Gemma 4 12B Q3: +8.55% Coding Performance From Tensor-Level Quantization Allocation

Model ReleasesDGX agent

Ive been experimenting with task-aware GGUF quants for months, taking inspiration from TASA and TAQO but pushing the allocation lower to to the tensor level. The basic idea is to generate a custom ima

Inverse-dynamics observer design for a linear single-track vehicle model with distributed tire dynamics

SafetyDGX agent

arXiv:2603.07499v3 Announce Type: replace-cross Abstract: Accurate estimation of the vehicle's sideslip angle and tire forces is essential for enhancing safety and handling performances in unknown dri

Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models

SafetyDGX agent

arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unk

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs

Model ReleasesDGX agent

arXiv:2608.11624v1 Announce Type: cross Abstract: Persuasion is a core dynamic of natural language communication, shaping how large language models (LLMs) update beliefs, resolve disagreements, and re

Market-Information-Aware Gated-LoRA of Foundation Models for Transferable Day-Ahead Electricity Price Forecasting

ResearchDGX agent

arXiv:2608.11359v1 Announce Type: new Abstract: Electricity price forecasting is crucial for market participants but remains difficult because prices are volatile, market-specific, and closely tied to

Mindgard raises $30M to handle security for AI models and applications

IndustryDGX agent

Mindgard Ltd., a leader in artificial intelligence cybersecurity, announced today that it raised 30 million in early funding to scale up its product in response to significant demand across the indust

Multi-Agent Embodied Autonomous Driving: From V2X Information Exchange to Shared World Models

SafetyDGX agent

arXiv:2606.13840v2 Announce Type: replace-cross Abstract: Autonomous driving is shifting from isolated vehicle intelligence toward multi-agent embodied systems that share perception, infer intent, and

Self-Harness: Harnesses That Improve Themselves

Model ReleasesDGX agent

arXiv:2606.09498v2 Announce Type: replace Abstract: The performance of LLM-based agents is jointly shaped by their base models and the harnesses that mediate their interaction with the environment. Be

Temperature-Driven Sequential Modeling for the Prediction of Annual Power Conversion Efficiency Profiles of Organic Photovoltaic Materials: Douala Case Study

ApplicationsDGX agent

arXiv:2608.11261v1 Announce Type: cross Abstract: Organic photovoltaic (OPV) materials are promising candidates for distributed solar energy in tropical regions, yet existing virtual screening tools r

VLM2Rec: Resolving Modality Collapse in Vision-Language Model Embedders for Multimodal Sequential Recommendation

ResearchDGX agent

arXiv:2603.17450v2 Announce Type: replace-cross Abstract: Sequential Recommendation (SR) in multimodal settings typically relies on small frozen pretrained encoders, which limits semantic capacity and

Who Thinks Best Depends on How Long You Let Them: Budget-Dependent Rankings in LLM Evaluation

ResearchDGX agent

arXiv:2608.12150v1 Announce Type: new Abstract: Standard evaluation of large language models assumes stable model rankings across inference conditions. We challenge this assumption by varying the toke

12 Aug 2026

An adaptive and evolvable deep reinforcement learning framework for weather prediction

Model ReleasesDGX agent

arXiv:2608.09948v1 Announce Type: cross Abstract: No single AI weather model excels at all variables, pressure levels, and lead times. Rather than building yet another architecture, we reframe the for

Compute-Optimal Is Not Cluster-Optimal: Systems-Aware Scaling for Sparse Mixture-of-Experts

ResearchDGX agent

arXiv:2608.10605v1 Announce Type: cross Abstract: In large-scale pretraining, the algorithm, architecture, and systems decisions are conventionally made in disconnected stages. A scaling law stage sel

Cost-Efficient Estimation of General Abilities Across Benchmarks

Model ReleasesDGX agent

arXiv:2604.01418v2 Announce Type: replace Abstract: Thousands of diverse benchmarks have been developed to measure the quality of large language models (LLMs). Yet prior work has demonstrated that LLM

DeepSeek V4 Pro 0813 (on OpenRouter)

Model ReleasesDGX agent

DeepSeek V4 Pro 0813 (on OpenRouter) The latest DeepSeek Pro model is now available, via API only. I had to link to OpenRouter because DeepSeek don't have any obvious announcement page for their new m

Do LLM Recommenders Know When They're Hallucinating? Auditing Confidence Calibration in Catalog Faithfulness

Model ReleasesDGX agent

arXiv:2608.10008v1 Announce Type: cross Abstract: LLM recommenders for top-K item suggestion regularly emit titles outside the target catalog. Prior audits measure this as a binary out-of-domain rate;

Expert-Guided g-computation with Large Language Models for Estimating Causal Effects on Timings: Applications to Hospital Quality Improvement

SafetyDGX agent

arXiv:2608.10339v1 Announce Type: cross Abstract: Hospital quality improvement (QI) programs routinely face multiple candidate interventions to optimize hospital flow, but existing methods struggle to

Introspective Attention Modulation for Safe Text-to-Image Generation

Model ReleasesDGX agent

arXiv:2607.14945v2 Announce Type: replace Abstract: State-of-the-art flow based text-to-image (T2I) models exhibit remarkable generative abilities but remain vulnerable to producing unsafe content. Pr

Leveraging Large Language Models for Causal Discovery: a Constraint-based, Argumentation-driven Approach

SafetyDGX agent

arXiv:2602.16481v2 Announce Type: replace Abstract: Causal discovery seeks to uncover causal relations from data, typically represented as causal graphs, and is essential for predicting the effects of

Modelling Geographic Atrophy Progression using Implicit Neural Representations

ResearchDGX agent

arXiv:2608.10807v1 Announce Type: cross Abstract: Age-related Macular Degeneration (AMD) is the major cause of blindness in the Western world. Its late dry phase is characterised by irreversible atrop

PEAK: Precise and Persistent Concept Erasure via k-Sparse Autoencoders

Model ReleasesDGX agent

arXiv:2608.10985v1 Announce Type: new Abstract: Erasing concepts from large-scale text-to-image (T2I) diffusion models has become increasingly crucial due to the growing concerns over copyright infrin

Persona Conditioning as an Assessor-Sensitivity Probe for LLM-Based IR Evaluation

Model ReleasesDGX agent

arXiv:2608.10385v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as relevance assessors in information retrieval (IR) evaluation, raising questions about how assess

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2608.11191v1 Announce Type: cross Abstract: GUI Visual Grounding is a fundamental capability for GUI agents. Existing models typically freeze their parameters after deployment, limiting their ab

11 Aug 2026

AirFlow: Context Preserving and Multi-Rate State Modeling for Air Quality Forecasting

ApplicationsDGX agent

arXiv:2608.09775v1 Announce Type: new Abstract: Accurate air quality forecasting is essential for public health and urban environmental management, but remains challenging because pollutant channels d

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

SafetyDGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

Can We Optimize the Performance-Carbon Emission Break-Even Point?: The Quest for Greener LLMs

Model ReleasesDGX agent

arXiv:2608.08744v1 Announce Type: cross Abstract: The carbon footprint of any deployed Large Language Model (LLM) accumulates during inference, where repeated use of the model substantially exceeds th

Damage Classification for 3D Point Cloud Data via 3D Data Analysis and Vision Foundation Model-based 2D Projections

ResearchDGX agent

arXiv:2608.08955v1 Announce Type: new Abstract: Fine-grained damage classification of 3D point cloud data (PCD) remains a persistent challenge, constrained by high computational demands and limited la

F2STNet: Fair and Federated Spectral-Temporal Modeling for Graph Forecasting

SafetyDGX agent

arXiv:2608.09082v1 Announce Type: new Abstract: Spatiotemporal prediction on graph-structured data is central to traffic forecasting and environmental monitoring, yet decentralized and heterogeneous d

Focus particles and scalar inferences across humans and language models

ResearchDGX agent

arXiv:2608.08227v1 Announce Type: new Abstract: Focus particles such as 'even' and 'only' are central to formal semantic theories that posit structured representations over sets of alternatives. 'Even

Foundation Models are Implicit Deepfake Detectors

ResearchDGX agent

arXiv:2608.09427v1 Announce Type: new Abstract: Pretrained self-supervised representations have emerged as a core component of current deepfake detection methods, yet it remains unclear which of their

From Objectives to What Models Learn: A Landau Theory of Invariant Learning

TutorialsDGX agent

arXiv:2608.09396v1 Announce Type: new Abstract: Invariant learning seeks representations that remain predictive across environments, yet the behavior of its objectives along the regularization path is

GeoPhysAdapter: Scale-Matched Geophysical Adaptation for Cross-Domain Landslide Mapping with Vision Foundation Models

Local AiDGX agent

arXiv:2608.09325v1 Announce Type: new Abstract: Newly triggered landslides rarely carry immediate annotations, so cross-domain transferability determines the value of landslide mapping for emergency r

GLocFM: A Geometry-Aware Foundation Model for 3D Indoor Wireless Localization

Local AiDGX agent

arXiv:2608.09285v1 Announce Type: cross Abstract: Learning-based wireless localizers often fail to utilize geometric information about the propagation environment, limiting their ability to exploit no

Introducing Unsloth Desktop app

Model ReleasesDGX agent

Hi LocalLlama, we're super excited to release Unsloth Desktop today! 🦥 It's the first desktop app that enables you to run and train models locally. Open-source. Available on Mac, Windows, and Linux Su

Learning Preference Adaptation for Large Language Model Personalization via Verbal Reinforcement Learning

SafetyDGX agent

arXiv:2608.09507v1 Announce Type: cross Abstract: Natural language user preferences provide an interpretable interface for LLM personalization. However, universal preference summaries often contain in

MiniMax-H3: ~38 GB less VRAM with Runtime LoRA Bypass — DoRA Dynamic LoRA Loader v1.0.39

Local AiDGX agent

GitHub: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader Release v1.0.39: https://github.com/xmarre/ComfyUI-DoRA-Dynamic-LoRA-Loader/releases/tag/v1.0.39 Also available through ComfyUI Manag

MiraMind: Benchmarking Reliable Mental Health Reasoning beyond Answer Accuracy

Model ReleasesDGX agent

arXiv:2512.09636v3 Announce Type: replace Abstract: Mental-health reasoning with large language models (LLMs) is an evidence-constrained judgment problem: models must transform limited, subjective, an

MotionCraft: Latent World Modeling with Sparse Attention for Visual Upscaling

SafetyDGX agent

arXiv:2608.08553v1 Announce Type: new Abstract: Video super-resolution (VSR) aims to recover high-fidelity high-resolution videos from low-resolution inputs and is central to applications ranging from

MRI super-resolution in ten sampling steps using a diffusion bridge model

ResearchDGX agent

arXiv:2608.08819v1 Announce Type: new Abstract: Objective. MRI provides excellent soft-tissue contrast, but long acquisition times can cause patient discomfort and lead to motion artifacts, forcing a

NVIDIA and Local AI Community Fuel Open Source Models and Intelligent Agents

Local AiDGX agent

The open source ecosystem is making it easier for AI enthusiasts and developers to build, customize and run increasingly capable agents locally. Throughout August, NVIDIA is celebrating the partners a

NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents

Model ReleasesDGX agent

NVIDIA Nemotron 3.5 Lightning is an open‑source mixture‑of‑experts language model totaling 30 B parameters with only 3 B active during inference, designed to serve high‑volume, low‑latency execution f

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, ev…

Model ReleasesDGX agent

Open source is so back. Zuck just announced Meta is opening the weights for Muse Glimmer, with Muse Spark 1.2 coming soon But a year ago, everyone doubted Meta's position in the AI race In an intervie

PluginEval: A Diagnostic Benchmark for Fine-Grained Error Attribution in Function Calling

Model ReleasesDGX agent

arXiv:2608.08700v1 Announce Type: new Abstract: Reliable evaluation of tool routing is critical as Large Language Models increasingly operate as autonomous agents. Current benchmarks face three struct

← Previous
1…225226227228229…1018
Next →