AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,428 results
Local Ai

Uncertainty-Aware Simulation-Based Inference for Operations Research with Large Language Models

DGX agent

arXiv:2608.00019v1 Announce Type: new Abstract: Deploying large language models (LLMs) for operations research (OR) tasks remains challenging because correctness depends on a coherent modeling process

local-aiarxiv-cs-lg
4 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Where did the ambiguity go? Examining how multimodal models interpret polysemous words

DGX agent

arXiv:2608.00410v1 Announce Type: cross Abstract: Human language is highly polysemous. Many common words (e.g., 'bank' or 'palm') carry several distinct meanings that shape what humans communicate and

researcharxiv-cs-cl
4 Aug 2026
Local Ai

Accelerated Random-Sweep Gibbs Sampling for Gaussian Graphical Models via Dual Normal Factor Graphs

DGX agent

arXiv:2607.28706v1 Announce Type: cross Abstract: We study the convergence properties of the random-sweep Gibbs sampler for Gaussian graphical models with a thin-membrane prior. We demonstrate that th

local-aiarxiv-cs-lg
3 Aug 2026
Model Releases

Combining Large Language Models and Symbolic Reasoning for Multi-Robot Temporal Planning through Explainable Knowledge Bases

DGX agent

arXiv:2502.19135v2 Announce Type: replace Abstract: We present PLANTOR, a framework for generating and executing multi-robot task plans from natural-language task descriptions through LLM-assisted kno

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

DreamQAS: Learning a Decision-Useful World Model for VQE-Efficient Quantum Architecture Search

DGX agent

arXiv:2607.29491v1 Announce Type: cross Abstract: Reinforcement-learning-based quantum architecture search (RL-QAS) repeatedly optimizes a variational quantum eigensolver (VQE) after extending a circu

safetyarxiv-cs-ai
3 Aug 2026
Safety

Learning Latent Reasoning Traces for Scalar Reward Models End-to-End

DGX agent

arXiv:2607.29185v1 Announce Type: new Abstract: Reward models (RMs) are central to aligning large language models with human preferences via reinforcement learning. Although traditional scalar RMs ena

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

M3-DuplexBench: A Multi-Turn, Multilingual, Multidomain Benchmark for Full-Duplex Spoken Dialogue Models

DGX agent

arXiv:2607.29125v1 Announce Type: new Abstract: Full-duplex spoken dialogue systems (FDSDSs) can listen while speaking, enabling natural behaviors such as smooth turn-taking, backchannel handling, and

model-releasesarxiv-cs-cl
3 Aug 2026
Model Releases

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also go…

DGX agent

📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also going open-weights to meet you all!🎉 Qwen3.8-Max, a new bar for

model-releasesqwen--x
3 Aug 2026
Model Releases

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

DGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

model-releasesarxiv-cs-ai
3 Aug 2026
Local Ai

Optimal Resource Allocation for ML Model Training and Deployment under Concept Drift

DGX agent

arXiv:2512.12816v2 Announce Type: replace Abstract: We study how to allocate resources for training and deployment of machine learning (ML) models under concept drift and limited budgets. We consider

local-aiarxiv-cs-lg
3 Aug 2026
Agents

SERUM: State Extraction and Refinement for User Modeling

DGX agent

arXiv:2607.29181v1 Announce Type: cross Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these mode

agentsarxiv-cs-ai
3 Aug 2026
Local Ai

For MoE models the arithmetic splits in two: capacity follows total params, speed follows active

DGX agent

A few people asked for this after the bandwidth thread, so here it is on its own instead of buried in a comment. The dense rule was simple: every token reads every weight, so tokens/sec ≈ bandwidth ÷

local-air-ollama
1 Aug 2026
Local Ai

Can Vision-Language Models Reason about AI Edits in Images?

DGX agent

arXiv:2607.28464v1 Announce Type: new Abstract: Detection and localization of AI-tampered images are critical for trustworthy AI, yet modern generative models have made such manipulations increasingly

local-aiarxiv-cs-cv
31 Jul 2026
Applications

Enhancing Irregular Time Series Forecasting with Continuous-Time Modeling Framework

DGX agent

arXiv:2607.28035v1 Announce Type: new Abstract: Irregular multivariate time series are widely encountered in applications such as healthcare monitoring, human activity recognition, and environmental s

applicationsarxiv-cs-lg
31 Jul 2026
Tutorials

Latent Matters: Learning Deep State-Space Models

DGX agent

arXiv:2602.23050v2 Announce Type: replace Abstract: Deep state-space models (DSSMs) enable temporal predictions by learning the underlying dynamics of observed sequence data. They are often trained by

tutorialsarxiv-cs-lg
31 Jul 2026
Research

Negative controls reveal volume-driven confounding in radiomics and imaging foundation model features

DGX agent

arXiv:2607.28423v1 Announce Type: new Abstract: Radiomics and imaging foundation models promise non-invasive biomarkers of tumour biology, yet predictive signatures may reflect tumour volume or acquis

researcharxiv-cs-cv
31 Jul 2026
Safety

The Kinetics of Training: A Driven-Nucleation Rate Law for Emergence, Plasticity Loss, and Circuit Control in Language Models

DGX agent

arXiv:2607.27281v1 Announce Type: new Abstract: A capability appears in a language model when the last parts of its circuit align in one stochastic attempt, and getting all but one right is worth noth

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

THGFM: Dual-Branch Temporal Heterogeneous Graph Fusion Model

DGX agent

arXiv:2607.27303v1 Announce Type: cross Abstract: Temporal heterogeneous graphs offer a natural abstraction for dynamic relational systems in which diverse node and relation types co-exist and evolve

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Using Large Language Models for Idea Generation in Innovation

DGX agent

arXiv:2607.27553v1 Announce Type: cross Abstract: This research evaluates the efficacy of large language models (LLMs) in generating new product ideas. To do so, we compare three pools of ideas for ne

model-releasesarxiv-cs-cl
31 Jul 2026
Tutorials

What Makes Graph Unified? Principles and Generative Sliding-Window Transformer for Graph Foundation Models

DGX agent

arXiv:2607.27966v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) have recently emerged as a promising paradigm for general-purpose graph learning, aiming to learn reusable knowledge that

tutorialsarxiv-cs-lg
31 Jul 2026
Local Ai

ActSWM: Action-Sensitive World Models for Long-Horizon Planning in Open-World Games

DGX agent

arXiv:2607.26712v1 Announce Type: new Abstract: Latent world models support efficient model-predictive control by optimizing future control sequences in latent space and replanning in a receding-horiz

local-aiarxiv-cs-ro
30 Jul 2026
Model Releases

ARC-Encoder: learning compressed text representations for large language models

DGX agent

arXiv:2510.20535v2 Announce Type: replace Abstract: Recent techniques such as retrieval-augmented generation or chain-of-thought reasoning have led to longer contexts and increased inference costs. Co

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Archetypes or ability? Clustering for modelling student mathematical competence

DGX agent

arXiv:2607.26063v1 Announce Type: cross Abstract: Personalised learning systems often assume that mathematical ability is combined of discrete abilities, acquired sequentially and dependent upon first

model-releasesarxiv-cs-lg
30 Jul 2026
Agents

From Conceptual Hydrologic Models to Conceptually Interpretable Neural Networks: A Snow-Water Mass-Conserving-Perceptron Framework for Discovering Catchment-Scale Precipitation-Storage-Runoff Representations

DGX agent

arXiv:2607.26492v1 Announce Type: new Abstract: The Mass-Conserving Perceptron (MCP) establishes a modeling paradigm in which conceptual hydrologic models can be reformulated as physically constrained

agentsarxiv-cs-lg
30 Jul 2026
Model Releases

HumanCLAW: Can Vision-Language Models Act Through a Body?

DGX agent

arXiv:2607.27180v1 Announce Type: new Abstract: Evaluating whether a vision-language model (VLM) can act through a physical body is challenging. The outcome of an action couples the VLM's decision wit

model-releasesarxiv-cs-cv
30 Jul 2026
Safety

OptimismBench: Forecasting Bias and the Alignment Effect in Language Model Judgment

DGX agent

arXiv:2607.26981v1 Announce Type: new Abstract: Large language models are increasingly used as decision aids whose probability judgments shape downstream choices. Whether those judgments carry a syste

safetyarxiv-cs-cl
30 Jul 2026
Research

Relation Geometry in Semantic Space of Language Models

DGX agent

arXiv:2607.26762v1 Announce Type: new Abstract: When it comes to generating vector representations of words, current language models are achieving high-quality results. However, what is not known is t

researcharxiv-cs-cl
30 Jul 2026
Research

Scientific Knowledge Discovery in the Age of Large Language Models

DGX agent

arXiv:2607.26670v1 Announce Type: cross Abstract: The rapid growth of scholarly literature has made identifying relevant publications increasingly difficult, and conventional search systems still depe

researcharxiv-cs-cl
30 Jul 2026
Safety

Thinking Under Uncertainty: Evidence Use and Information-Seeking in Language Models

DGX agent

arXiv:2607.26845v1 Announce Type: new Abstract: Inference-time thinking improves the performance of large language models, but aggregate outcomes do not reveal whether models use available evidence mo

safetyarxiv-cs-lg
30 Jul 2026
Model Releases

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT…

DGX agent

We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a f

model-releasesopenai--x
30 Jul 2026
Model Releases

CogArena: A Multimethod Evaluation of Cognitive Ability Structure in Large Language Models

DGX agent

arXiv:2607.24999v1 Announce Type: cross Abstract: LLM cognitive scores are increasingly summarized as per-ability profiles whose dimensions should converge across tasks, respond selectively to matched

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CoTinyVLA: Chain-of-Thought Distillation for a Sub-Billion-Parameter Vision-Language-Action Model

DGX agent

arXiv:2607.25487v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models translate natural-language commands into robot action sequences, but leading systems on the LIBERO-Plus robustness b

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

From CUDA to MLX: How K-Search Brings Decades of Kernel Expertise to Apple Silicon

DGX agent

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruction. We face a new

model-releasesberkeley-ai-research
29 Jul 2026
Model Releases

How Small Can You Go? A Controlled Study of LoRA Rank, Target Modules, and Quantization Trade-offs for Text-to-SQL on a 60M-Parameter Model

DGX agent

arXiv:2607.25583v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) and low-bit quantization are now standard tools for adapting language models under tight compute budgets, yet the

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Mage-VL: An Efficient Codec-Native Streaming Multimodal Foundation Model

DGX agent

arXiv:2607.24904v1 Announce Type: cross Abstract: Standard vision-language models (VLMs) suffer from Moravec's paradox: they excel at complex offline visual reasoning but struggle with simple streamin

model-releasesarxiv-cs-cl
29 Jul 2026
Research

Raven: High-Recall Sequence Modeling with Sparse Memory Routing

DGX agent

arXiv:2607.25357v1 Announce Type: cross Abstract: Long-context recall in linear-time sequence models highlights a tradeoff in how they write to memory. State-based linear models, such as state-space m

researcharxiv-cs-ai
29 Jul 2026
Model Releases

Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT

DGX agent

arXiv:2607.25063v1 Announce Type: new Abstract: Developers judge a model checkpoint by how it behaves. After supervised fine-tuning (SFT), two checkpoints that perform about the same across relevant b

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Understanding the Cognitive Habits of Large Reasoning Models

DGX agent

arXiv:2506.21571v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs), which autonomously produce a reasoning Chain of Thought (CoT) before producing final responses, offer a promisi

model-releasesarxiv-cs-ai
29 Jul 2026
Research

Action from Adjacent Set in Physical Space Outperforms the Best Prediction in World Models

DGX agent

arXiv:2607.23602v1 Announce Type: cross Abstract: Controllers based on sampling and latent world models assign a predicted terminal cost to each candidate action sequence, choose the minimum, execute

researcharxiv-cs-ai
28 Jul 2026
Research

Autoregressive One-Step Generative Modeling for Dynamical System Forecasting

DGX agent

arXiv:2605.05540v2 Announce Type: replace Abstract: Fast surrogate modeling for high-dimensional physical dynamics requires more than low short-term error: useful models must roll out efficiently whil

researcharxiv-cs-lg
28 Jul 2026
Local Ai

CONSISTRE: A Unified Consistency-Aware Framework for Document-Level Relation Extraction with Large Language Models

DGX agent

arXiv:2607.24312v1 Announce Type: new Abstract: Document-level relation extraction (DocRE) aims to extract relations among multiple entities across extended contexts while maintaining consistency acro

local-aiarxiv-cs-cl
28 Jul 2026
Model Releases

ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams

DGX agent

arXiv:2607.24707v1 Announce Type: new Abstract: Entity-Relationship Diagrams (ERDs) are central to conceptual database design, yet they are typically available only as rendered images rather than mach

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias

DGX agent

arXiv:2607.22837v1 Announce Type: cross Abstract: Many organizations aim to adapt language models for internal use, both to improve performance on domain-specific tasks and to address privacy concerns

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

Hierarchical Reinforcement Learning with Optimal Level Synchronization Based on Flow-Based Deep Generative Model

DGX agent

arXiv:2107.08183v2 Announce Type: replace Abstract: High-dimensional state and action spaces combined with sparse reward structures in reinforcement learning (RL) environments typically require advanc

model-releasesarxiv-cs-lg
28 Jul 2026
Tutorials

Like a bilingual baby: The advantage of visually grounding a bilingual language model

DGX agent

arXiv:2210.05487v3 Announce Type: replace Abstract: Unlike most neural language models, humans learn language in a rich, multi-sensory and, often, multi-lingual environment. Current language models ty

tutorialsarxiv-cs-cl
28 Jul 2026
Research

Moving-Horizon Estimation and Nonlinear Model Predictive Control of Cable-Driven Soft Manipulators

DGX agent

arXiv:2607.24029v1 Announce Type: new Abstract: Precise control of soft manipulators remains challenging due to the difficulty of developing accurate yet computationally tractable models for model-bas

researcharxiv-cs-ro
28 Jul 2026
Local Ai

Real-Time Human-Centric World Modeling for Upper-Body Human-Object Interaction

DGX agent

arXiv:2607.23517v1 Announce Type: new Abstract: We present a real-time human-centric world model for upper-body interactive generation, aiming to synthesize coherent local world dynamics centered on a

local-aiarxiv-cs-cv
28 Jul 2026
Safety

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation

DGX agent

arXiv:2607.23125v1 Announce Type: new Abstract: Post-training enables vision-language models (VLMs) to understand human instructions and perform various downstream tasks. Current post-training methods

safetyarxiv-cs-lg
28 Jul 2026
← Previous
1…7677787980…1259
Next →