AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlog
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
Model Releases

Limits to scalable evaluation at the frontier: LLM as Judge won't beat twice the data

DGX agent

arXiv:2410.13341v4 Announce Type: replace Abstract: High quality annotations are increasingly a bottleneck in the explosively growing machine learning ecosystem. Scalable evaluation methods that avoid

model-releasesarxiv-cs-lg
18 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

MathGen: Revealing the Illusion of Mathematical Competence through Text-to-Image Generation

DGX agent

arXiv:2603.27959v3 Announce Type: replace Abstract: Modern generative models have demonstrated the ability to solve challenging mathematical problems. In many real-world settings, however, mathematica

model-releasesarxiv-cs-cv
18 Aug 2026
Research

Quantum Models with Multi-Stage Training for Compositional Concept Generalization

DGX agent

arXiv:2608.15601v1 Announce Type: new Abstract: Compositional Concept Generalization (CoCoGen), the ability to systematically recombine learned primitives in novel contexts, is a key challenge for mul

researcharxiv-cs-lg
18 Aug 2026
Local Ai

Rotate Disks to Reach Farther: Design and Modeling of a Novel Reconfigurable Tendon Driven Manipulator

DGX agent

arXiv:2608.15946v1 Announce Type: new Abstract: Rerouting the tendon path in tendon driven continuum manipulators (TDCMs) enables a broad range of deformation modes. This work presents a Reconfigurabl

local-aiarxiv-cs-ro
18 Aug 2026
Agents

The Hallucination Snowball: Modeling Error Propagation as State Transitions in Multi-Agent LLM Pipelines

DGX agent

arXiv:2608.14588v1 Announce Type: new Abstract: Sequential multi-agent LLM pipelines chain specialized agents without verification at handoffs, creating a structural flaw with measurable and severe co

agentsarxiv-cs-ai
18 Aug 2026
Applications

Time-Aware Validation of Machine Learning Fuel Consumption Models: Evidence from 1,Hz Operational Data, CCGS extit{Sir Wilfrid Laurier}

DGX agent

arXiv:2608.16833v1 Announce Type: new Abstract: Ship fuel consumption (SFC) prediction supports vessel operation optimisation, emissions estimation, and decision support systems (DSS) for sustainable

applicationsarxiv-cs-lg
18 Aug 2026
Applications

Towards Unified Approaches in Self-Supervised Event Stream Modeling: Progress and Prospects

DGX agent

arXiv:2502.04899v3 Announce Type: replace-cross Abstract: The proliferation of digital interactions across diverse domains, such as healthcare, e-commerce, gaming, and finance, has resulted in the gen

applicationsarxiv-cs-ai
18 Aug 2026
Research

Translating finite-domain integer constraint models to CP/SMT/ILP/PB/SAT solvers with CPMpy

DGX agent

arXiv:2608.15143v1 Announce Type: new Abstract: Constraint solving is a declarative approach for solving combinatorial satisfaction and optimization problems. The user specifies their problem through

researcharxiv-cs-ai
18 Aug 2026
Research

Vision-Based Calorie Estimation for Bangladeshi Street Food: A Comparative Study of Detection and Regression Models

DGX agent

arXiv:2509.01415v2 Announce Type: replace Abstract: With obesity emerging as a major global health concern, accurate calorie estimation systems have become increasingly important for effective dietary

researcharxiv-cs-cv
18 Aug 2026
Model Releases

Whose Gold? Annotator-Pool Disagreement Is Large at the Item Level, and Hidden by Small Leaderboards

DGX agent

arXiv:2608.15980v1 Announce Type: new Abstract: Preference benchmarks are built by hiring annotators, and the identity of those annotators is treated as an implementation detail. We measure what that

model-releasesarxiv-cs-cl
18 Aug 2026
Safety

Consistent Model Chasing Is Minimax Optimal: The Exact Value of Scalar Adversarial Adaptive Control under Large Parametric Uncertainty

DGX agent

arXiv:2608.13651v1 Announce Type: cross Abstract: We solve exactly a fundamental problem of adaptive control against adversarial disturbances: regulate the scalar system x_{t+1} = ax_t + u_t + w_t, x_

safetyarxiv-cs-lg
17 Aug 2026
Research

Cross-Disciplinary Taxonomy and Modeling of Misunderstanding Generation, Amplification, and Detection, from Pragmatics to AI Agents

DGX agent

arXiv:2608.13604v1 Announce Type: new Abstract: Detection of misunderstanding is an urgent problem to solve because communication has moved away from real-time, in-person interaction and is increasing

researcharxiv-cs-ai
17 Aug 2026
Model Releases

EXL3 seems to be fading from the r/LocalLLaMa consciousness, and while I suspected it, I'm surprised at this point in time.

DGX agent

EXL3 is an alternative to llama.cpp. And while there is extensive tooling for llama.cpp, EXL3's primary deployment (TabbyAPI), has a OpenAI compatible API so it shouldn't matter. Why won't this tool m

model-releasesr-localllama
17 Aug 2026
Research

Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning

DGX agent

arXiv:2608.14290v1 Announce Type: new Abstract: We introduce Mobius-v0, an architecture that comprises a globally shared Memory (FFN) that stores knowledge vectors and multiple Reasoners (Self-Attn) t

researcharxiv-cs-ai
17 Aug 2026
Research

Multiphase-Diff: Diffusion-Based Generative Modeling for High-Contrast Multiphase Physical Systems with Sharp Interfaces

DGX agent

arXiv:2608.13669v1 Announce Type: new Abstract: Physics-constrained diffusion for high-contrast, sharp-interface multiphase fields faces three coupled difficulties. At coefficient jumps, expanded poin

researcharxiv-cs-cv
17 Aug 2026
Model Releases

PRM-as-a-Judge 1.5: A Toolkit for Robot Process Assessment

DGX agent

arXiv:2608.14284v1 Announce Type: cross Abstract: Fine-grained robotic evaluation matters for understanding embodied models, going beyond binary success rates and rule-based process scores. We present

model-releasesarxiv-cs-cv
17 Aug 2026
Research

Scaling Creative Writing Beyond Story-Centric Data with Attribute-Guided Genre Expansion

DGX agent

arXiv:2608.13947v1 Announce Type: new Abstract: High-quality creative writing data for large language models (LLMs) remains dominated by story-centric data, limiting models' ability to follow the stru

researcharxiv-cs-cl
17 Aug 2026
Research

Secret-Stego Dissimilarity as a Design Axis: Invertible Coverless Image Steganography with Diffusion Models

DGX agent

arXiv:2608.13597v1 Announce Type: cross Abstract: Coverless image steganography (CIS) synthesizes a stego image rather than modifying an existing cover image, enabling authorized recipients to reconst

researcharxiv-cs-ai
17 Aug 2026
Model Releases

Based on an accelerating frontier -> local trajectory, expect a ~30b param 'Mythos at home' by as soon as Jan 2027 (rationalisation below)

DGX agent

Including the rationalisation for the data below - this is a more robust version of an earlier post I did similar to this - explaining below: How I chose the comparisons The basic question I’m trying

model-releasesr-localllama
16 Aug 2026
Model Releases

Ollama works locally, but my coding-agent integration does not

DGX agent

I’m running Ollama on an M4 Pro Mac with 48 GB unified memory and testing local coding models for repository work. Direct Ollama inference is fine. Qwen3.8 27B-MLX runs well enough for me, and the loc

model-releasesr-ollama
16 Aug 2026
Model Releases

Beyond Correctness: Benchmarking and Aligning Response Behaviors in Hybrid-Thinking MLLMs

DGX agent

arXiv:2608.12781v1 Announce Type: new Abstract: Hybrid-thinking multimodal large language models (MLLMs) allow a single model to alternate between deliberative thinking and latency-efficient non-think

model-releasesarxiv-cs-cv
14 Aug 2026
Model Releases

Do LLMs Beat Nash? Testing Decentralized Coordination in Self-Play Multi-Agent Games

DGX agent

arXiv:2608.12547v1 Announce Type: cross Abstract: Large language model agents deployed without a central controller are often assumed to require communication to coordinate their actions. We ask what

model-releasesarxiv-cs-ro
14 Aug 2026
Model Releases

Edit2TikZ: A Comprehensive and Challenging Benchmark for Scientific Figure Editing with TikZ

DGX agent

arXiv:2608.13441v1 Announce Type: new Abstract: Although multimodal large language models (MLLMs) have shown substantial potential in visual understanding and graphic code generation, editing scientif

model-releasesarxiv-cs-cv
14 Aug 2026
Agents

Foam-Agent: A Large Language Model-Based Multi-Agent Framework for Automating Computational Fluid Dynamics Workflows

DGX agent

arXiv:2505.04997v3 Announce Type: replace Abstract: Computational fluid dynamics (CFD) has been the main workhorse of computational physics, yet its steep learning curve and fragmented, multi-stage wo

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

From Atomic Evidence to Logical Composition: Structured Compositional Reasoning over Compound Answer Options

DGX agent

arXiv:2608.12836v1 Announce Type: cross Abstract: Large language models often fail when answer options require combining atomic judgments under explicit logical operators, even when they judge the ind

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

How Do VLMs Behave When Blind or Misled? Behavioral Evaluation of VLMs on Scientific Figures

DGX agent

arXiv:2608.13267v1 Announce Type: cross Abstract: Existing vision-language model (VLM) benchmarks emphasize perception and reasoning accuracy (how well VLMs describe and reason about what they see in

model-releasesarxiv-cs-ai
14 Aug 2026
Model Releases

LoRA-Diffusion: Parameter-Efficient Fine-Tuning via Low-Rank Trajectory Decomposition

DGX agent

arXiv:2608.12328v1 Announce Type: new Abstract: Parameter-efficient fine-tuning methods such as LoRA have transformed the adaptation of large autoregressive language models, enabling task-specific cus

model-releasesarxiv-cs-cl
14 Aug 2026
Local Ai

Poll results 6 months later: When will we have Opus level with 30b model? Optimists win!

DGX agent

Poll at beginning of the year: https://www.reddit.com/r/LocalLLaMA/comments/1qj935h/poll_when_will_we_have_a_30b_open_weight_model_as/ The least voted option, 6 months, wins in my opinion, with 18 vot

local-air-localllama
14 Aug 2026
Model Releases

Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs

DGX agent

arXiv:2608.12675v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) is widely used to improve the performance of Large Language Models (LLMs) in answering user queries. Existing priva

model-releasesarxiv-cs-ai
14 Aug 2026
Research

QuISE: Defense against Typographic Attacks on VLMs via Query-Irrelevant Semantic Editing

DGX agent

arXiv:2608.13119v1 Announce Type: new Abstract: Typographic attacks pose a critical threat to vision-language models (VLMs) by injecting misleading text into images and causing models to rely on adver

researcharxiv-cs-cv
14 Aug 2026
Model Releases

QuoteBench: How Matched Scores Can Hide Command-Path Failures

DGX agent

arXiv:2608.13547v1 Announce Type: new Abstract: LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot disti

model-releasesarxiv-cs-ai
14 Aug 2026
Applications

Towards Socially Compliant Navigation in Deep Reinforcement Learning via Proxemics-Based Reward Modeling

DGX agent

arXiv:2608.12917v1 Announce Type: new Abstract: Developing effective robot navigation methods in crowded environments is essential for real-world applications. Although recent deep reinforcement learn

applicationsarxiv-cs-lg
14 Aug 2026
Model Releases

Virtual Temperature Sensors in Power Transformers Using Neural Ordinary Differential Equations

DGX agent

arXiv:2608.13260v1 Announce Type: new Abstract: Accurate modeling and forecasting of power transformer thermal behavior are critical for reliability, asset lifetime, and optimized power system operati

model-releasesarxiv-cs-lg
14 Aug 2026
Model Releases

Why Do AI Agents Break Rules? How Framing, Context, and Social Signals Shape Compliance

DGX agent

arXiv:2608.12323v1 Announce Type: cross Abstract: Specifying a penalty can paradoxically convert a legal obligation into a cost-benefit calculation that favors violation. We demonstrate that this enfo

model-releasesarxiv-cs-ai
14 Aug 2026
Local Ai

From Self-Normal-Positioning to Omni-Directional Tracking: Real-Time Surface Modeling Enabled Probe Tilt Control for Robotic Ultrasound Imaging

DGX agent

arXiv:2608.11409v1 Announce Type: new Abstract: Ultrasound (US) provides real-time, radiation-free imaging, but the image quality depends strongly on how the probe is oriented against the patient body

local-aiarxiv-cs-ro
13 Aug 2026
Model Releases

Gemma 4 12B Q3: +8.55% Coding Performance From Tensor-Level Quantization Allocation

DGX agent

Ive been experimenting with task-aware GGUF quants for months, taking inspiration from TASA and TAQO but pushing the allocation lower to to the tensor level. The basic idea is to generate a custom ima

model-releasesr-localllama
13 Aug 2026
Safety

Inverse-dynamics observer design for a linear single-track vehicle model with distributed tire dynamics

DGX agent

arXiv:2603.07499v3 Announce Type: replace-cross Abstract: Accurate estimation of the vehicle's sideslip angle and tire forces is essential for enhancing safety and handling performances in unknown dri

safetyarxiv-cs-ro
13 Aug 2026
Safety

Is Convergence Inevitable? Tracing Output Homogeneity Back to Base Models

DGX agent

arXiv:2608.11426v1 Announce Type: new Abstract: The lack of diversity in LM content is widely attributed to the alignment process, but how and where exactly in the pipeline this collapse begins is unk

safetyarxiv-cs-cl
13 Aug 2026
Model Releases

Learning to Persuade Exposes How Easily LLMs Abandon Correct Beliefs

DGX agent

arXiv:2608.11624v1 Announce Type: cross Abstract: Persuasion is a core dynamic of natural language communication, shaping how large language models (LLMs) update beliefs, resolve disagreements, and re

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Market-Information-Aware Gated-LoRA of Foundation Models for Transferable Day-Ahead Electricity Price Forecasting

DGX agent

arXiv:2608.11359v1 Announce Type: new Abstract: Electricity price forecasting is crucial for market participants but remains difficult because prices are volatile, market-specific, and closely tied to

researcharxiv-cs-lg
13 Aug 2026
Industry

Mindgard raises $30M to handle security for AI models and applications

DGX agent

Mindgard Ltd., a leader in artificial intelligence cybersecurity, announced today that it raised 30 million in early funding to scale up its product in response to significant demand across the indust

industrysiliconangle
13 Aug 2026
Safety

Multi-Agent Embodied Autonomous Driving: From V2X Information Exchange to Shared World Models

DGX agent

arXiv:2606.13840v2 Announce Type: replace-cross Abstract: Autonomous driving is shifting from isolated vehicle intelligence toward multi-agent embodied systems that share perception, infer intent, and

safetyarxiv-cs-cv
13 Aug 2026
Model Releases

Self-Harness: Harnesses That Improve Themselves

DGX agent

arXiv:2606.09498v2 Announce Type: replace Abstract: The performance of LLM-based agents is jointly shaped by their base models and the harnesses that mediate their interaction with the environment. Be

model-releasesarxiv-cs-cl
13 Aug 2026
Applications

Temperature-Driven Sequential Modeling for the Prediction of Annual Power Conversion Efficiency Profiles of Organic Photovoltaic Materials: Douala Case Study

DGX agent

arXiv:2608.11261v1 Announce Type: cross Abstract: Organic photovoltaic (OPV) materials are promising candidates for distributed solar energy in tropical regions, yet existing virtual screening tools r

applicationsarxiv-cs-lg
13 Aug 2026
Research

VLM2Rec: Resolving Modality Collapse in Vision-Language Model Embedders for Multimodal Sequential Recommendation

DGX agent

arXiv:2603.17450v2 Announce Type: replace-cross Abstract: Sequential Recommendation (SR) in multimodal settings typically relies on small frozen pretrained encoders, which limits semantic capacity and

researcharxiv-cs-ai
13 Aug 2026
Research

Who Thinks Best Depends on How Long You Let Them: Budget-Dependent Rankings in LLM Evaluation

DGX agent

arXiv:2608.12150v1 Announce Type: new Abstract: Standard evaluation of large language models assumes stable model rankings across inference conditions. We challenge this assumption by varying the toke

researcharxiv-cs-ai
13 Aug 2026
Model Releases

An adaptive and evolvable deep reinforcement learning framework for weather prediction

DGX agent

arXiv:2608.09948v1 Announce Type: cross Abstract: No single AI weather model excels at all variables, pressure levels, and lead times. Rather than building yet another architecture, we reframe the for

model-releasesarxiv-cs-lg
12 Aug 2026
Research

Compute-Optimal Is Not Cluster-Optimal: Systems-Aware Scaling for Sparse Mixture-of-Experts

DGX agent

arXiv:2608.10605v1 Announce Type: cross Abstract: In large-scale pretraining, the algorithm, architecture, and systems decisions are conventionally made in disconnected stages. A scaling law stage sel

researcharxiv-cs-ai
12 Aug 2026
← Previous
1…286287288289290…1294
Next →