AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
Model Releases

Reasoning as Attractor Dynamics: Latent Memory Retrieval via Gibbs-Weighted Energy Minimization

DGX agent

arXiv:2606.24543v1 Announce Type: new Abstract: Large Language Models (LLMs) are traditionally viewed as autoregressive generators. However, from the perspective of collective computation, they functi

model-releasesarxiv-cs-lg
24 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Towards Spec Learning: Inference-Time Alignment from Preference Pairs

DGX agent

arXiv:2606.24004v1 Announce Type: cross Abstract: Steering a large language model (LLM) toward a desired behavior typically relies on an iterative process of hand-crafting a prompt based on a careful

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

DGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

ZONOS2 Technical Report

DGX agent

arXiv:2606.24320v1 Announce Type: cross Abstract: We present ZONOS2 8B, our latest TTS model, which achieves state-of-the-art naturalness, prosody, and voice cloning fidelity. We improve upon Zonos-v0

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Verifiable Search Is Not a Learnable Chain-of-Thought

DGX agent

arXiv:2606.21884v1 Announce Type: new Abstract: It is tempting to assume any task solvable by a short program can be taught to a model as its chain-of-thought: write the steps out, fine-tune, and the

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ASCII Art Turns LLMs into VLA Controllers

DGX agent

arXiv:2606.21470v1 Announce Type: cross Abstract: Vision--Language--Action (VLA) controllers are often built by extending vision--language models (VLMs) with action supervision, relying on multimodal

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies

DGX agent

arXiv:2606.20599v1 Announce Type: cross Abstract: Tree of Thought (ToT) search has become a promising direction for improving the reasoning capabilities of large language models, but deploying these m

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond 'One Language, One Script': Quantifying Orthographic Bias in Multilingual VLMs with PuMVR

DGX agent

arXiv:2606.20770v1 Announce Type: cross Abstract: Current Vision-Language Models (VLMs) are celebrated for their multilingual capabilities, yet they operate under a flawed assumption: that one languag

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Chehre: An Emoji-Prompted Video Dataset for Perceptually Diverse Facial Expression Recognition

DGX agent

arXiv:2606.21657v1 Announce Type: new Abstract: Facial expressions are nonverbal social signals used in human interaction, but facial expression recognition datasets often focus on static images, basi

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Chem2Gen-Bench: Benchmarking Chemical-to-Genetic Translation in Perturbation Response Space

DGX agent

arXiv:2606.21109v1 Announce Type: new Abstract: Virtual-cell and perturbation models are increasingly used to predict cellular responses for biomedical discovery, but chemical and genetic perturbation

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Confidently Wrong: Severity-Aware Calibration of Prompt-Injection Detectors under Attack Shift

DGX agent

arXiv:2606.22659v1 Announce Type: cross Abstract: Prompt-injection detectors are deployed as guards: a model scores an input and a downstream system trusts or blocks it on that score. I study the conf

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Double-Diffusion: Balancing Speed, Accuracy, and Uncertainty in Probabilistic Forecasting for Urban Sensor Networks

DGX agent

arXiv:2506.23053v3 Announce Type: replace Abstract: Urban sensor networks need forecasts that are accurate, carry useful uncertainty, and refresh fast enough to act on as new readings arrive. These go

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion

DGX agent

arXiv:2606.20924v1 Announce Type: new Abstract: In multi-object text-to-image (T2I) diffusion, ensuring semantic consistency between textual prompts and generated visual content is crucial for image s

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Fara-1.5: Scalable Learning Environments for Computer Use Agents

DGX agent

arXiv:2606.20785v1 Announce Type: cross Abstract: Collecting computer use data from human demonstrations is expensive and slow, motivating the need for scalable generation strategies. This requires tw

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Good-Enough LLM Obfuscation (GELO)

DGX agent

arXiv:2603.05035v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly served on shared accelerators where an adversary with read access to device memory can observe K

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation

DGX agent

arXiv:2602.00032v3 Announce Type: replace-cross Abstract: Synthetic faces from text-to-image (T2I) models pervade digital media, yet their demographic biases under emotionally conditioned prompts rema

safetyarxiv-cs-cv
23 Jun 2026
Tutorials

Keep The Essentials: Efficient Reference Conditioned Generation via Token Dropping

DGX agent

arXiv:2606.23682v1 Announce Type: new Abstract: Reference-based diffusion models enable highly controllable image generation by leveraging elements from input images to guide prompt-driven synthesis.

tutorialsarxiv-cs-cv
23 Jun 2026
Model Releases

LAYUP: Asynchronous decentralized gradient descent with LAYer-wise UPdates

DGX agent

arXiv:2410.05985v4 Announce Type: replace Abstract: The increasing size of deep learning models has made distributed training across multiple devices essential. Synchronous, centralized methods incur

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

Local Causal Attribution of Chain-of-Thought Reasoning

DGX agent

arXiv:2606.21821v1 Announce Type: new Abstract: Understanding the causal structure of a language model's thought process is a problem of significant importance for both transparency and safety. In thi

local-aiarxiv-cs-lg
23 Jun 2026
Model Releases

MMGist: A Comprehensive Multimodal Benchmark for 2027

DGX agent

arXiv:2606.22437v1 Announce Type: new Abstract: We conduct a systematic study of 18 widely used vision-language benchmarks and identify three major issues: 1) many items do not rely on visual cues and

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Predictions as Surrogates: Revisiting Surrogate Outcomes in the Age of AI

DGX agent

arXiv:2501.09731v2 Announce Type: replace-cross Abstract: We establish a formal connection between the decades-old surrogate outcome model in biostatistics and economics and the emerging field of pred

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Probabilistic Retrofitting of Learned Simulators

DGX agent

arXiv:2603.01949v2 Announce Type: replace Abstract: Dominant approaches for modelling Partial Differential Equations (PDEs) rely on deterministic predictions, yet many physical systems of interest are

researcharxiv-cs-lg
23 Jun 2026
Model Releases

ReNIO: Reweighting Negative Trajectory Importance for LLM On-Policy Distillation

DGX agent

arXiv:2606.23104v1 Announce Type: new Abstract: On-policy distillation (OPD) improves LLM reasoning by training a student model on its own generated outputs, but standard OPD treats all student-genera

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving

DGX agent

arXiv:2606.22840v1 Announce Type: new Abstract: We present RLM-Cascade, a proxy-layer system that applies speculative decoding at the response level to reduce LLM API costs without requiring model arc

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing

DGX agent

arXiv:2606.18774v2 Announce Type: replace Abstract: We present RouteJudge, an online pairwise preference evaluation framework for LLM routing systems, with a public platform available at https://route

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Sequential Minimal Optimization Algorithm for One-Class Support Vector Machines With Privileged Information

DGX agent

arXiv:2606.22210v1 Announce Type: new Abstract: One of the powerful techniques in data modeling is accounting for features that are available at the training stage, but are not available when the trai

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Set-based v.s. Distribution-based Representations of Epistemic Uncertainty: A Comparative Study

DGX agent

arXiv:2602.22747v2 Announce Type: replace Abstract: Epistemic uncertainty in neural networks is commonly modeled using two second-order paradigms: distribution-based representations, which rely on pos

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

TeleStyle V2: Beyond Content-Preserving Style Transfer with Self-Distillation and Distribution-Matching-Distillation

DGX agent

arXiv:2606.20709v1 Announce Type: new Abstract: Given a content reference and a style reference, content-preserving style transfer requires the model to generate stylized outputs with content and styl

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Towards Error-Free Long Video Generation

DGX agent

arXiv:2606.22370v1 Announce Type: new Abstract: Recent advances in video generation have made minute-level synthesis possible; however, generating long videos remains challenging due to error accumula

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Understanding Parallel Samplers in Masked Diffusion via Random Walks on Graphs

DGX agent

arXiv:2606.22976v1 Announce Type: new Abstract: In this paper, we propose using random walks on graphs as a verifiable sandbox to study different parallel sampling strategies in masked diffusion model

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

UniRank: Unified Rank Allocation for Low-Rank LLM Compression

DGX agent

arXiv:2606.21847v1 Announce Type: new Abstract: Low-rank decomposition serves as a promising compression paradigm for large language models, however, rank allocation remains challenging: manual rules

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

Human intelligence is fundamentally a collective intelligence. We solve complex problems by participating in a vast cultural network that bu…

DGX agent

Human intelligence is fundamentally a collective intelligence. We solve complex problems by participating in a vast cultural network that builds upon ideas across generations. I believe the strongest

agentsdavid-ha--x
22 Jun 2026
Model Releases

Try it out! We are seeing amazing results with GLM 5.2!

DGX agent

Try it out! We are seeing amazing results with GLM 5.2! it is indeed quite good! don't try it in claude code/codex - those harnesses are overly tuned for their proprietary models dcode (deepagents cod

model-releasesharrison-chase--x
19 Jun 2026
Model Releases

A Lightweight Multi-Agent Framework for Automated Concrete Barrier Design

DGX agent

arXiv:2606.12040v1 Announce Type: new Abstract: The design of reinforced concrete highway barriers is a safety-critical process that requires strict compliance with regulatory provisions such as the A

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Categorical Prior Lock-in: Why In-Context Learning Fails for Structured Data

DGX agent

arXiv:2606.11961v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as conditional generators for structured data, relying on in-context learning (ICL) to adapt to new

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

From Consumption to Reflection: Designing Human-AI Relations for Stable Reasoning

DGX agent

arXiv:2606.11195v1 Announce Type: cross Abstract: Large language models (LLMs) have transformed how humans access information, but not how we reason with it. Their fluency accelerates consumption whil

safetyarxiv-cs-ai
11 Jun 2026
Hardware

INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration

DGX agent

arXiv:2606.11440v1 Announce Type: new Abstract: Existing multi-agent LLM orchestration methods, ranging from brute-force ensembles to learned routers, select models and topologies based on task and mo

hardwarearxiv-cs-ai
11 Jun 2026
Model Releases

Neural-Parameterized Cellular Automata for Wildfire Spread

DGX agent

arXiv:2606.11676v1 Announce Type: cross Abstract: Traditional wildfire models rely on rigid, low-dimensional parameters and static fuel maps, frequently underpredicting fire spread. To address this we

model-releasesarxiv-cs-lg
11 Jun 2026
Research

On Subquadratic Architectures: From Applications to Principles

DGX agent

arXiv:2606.12364v1 Announce Type: new Abstract: Transformers dominate modern sequence modeling, but their quadratic attention incurs substantial computational cost. Subquadratic architectures offer a

researcharxiv-cs-lg
11 Jun 2026
Model Releases

RankVR: Low-Rank Structure Perception and Value Recalibration for Robust Composed Image Retrieval

DGX agent

arXiv:2606.11689v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) constitutes a pivotal paradigm requiring models to perform joint reasoning on reference images and modification texts. Ho

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Running Gemma 4 QAT 12B on an 8GB GPU at 16k context — measured the KV-cache tradeoffs

DGX agent

This post discusses running Google's Gemma 4 QAT (Quantized Aware Training) 12B model on a GPU with 8GB of memory while maintaining a 16k token context window. The author likely shares performance ben

model-releasesr-ollama
11 Jun 2026
Model Releases

SPEAR: A System for Post-Quantization Error-Adaptive Recovery Enabling Efficient Low-Bit LLM Serving

DGX agent

arXiv:2606.11244v1 Announce Type: cross Abstract: Efficient large language model (LLM) serving is increasingly constrained by deployment cost. Quantization is a key technique for reducing serving cost

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

STEAM: Squeeze and Transform Enhanced Attention Module

DGX agent

arXiv:2412.09023v3 Announce Type: replace Abstract: Channel and spatial attention mechanisms introduced in earlier work enhance the representational capabilities of deep convolutional neural networks

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

TAHOE: Text-to-SQL with Automated Hint Optimization from Experience

DGX agent

arXiv:2606.12387v1 Announce Type: cross Abstract: Large Language Models (LLMs) have democratized database access through Text-to-SQL, but moving from prototypes to production remains difficult. Real d

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Unifying Learning Dynamics and Generalization in Transformers Scaling Law

DGX agent

arXiv:2512.22088v3 Announce Type: replace-cross Abstract: The scaling law, a cornerstone of Large Language Model (LLM) development, predicts improvements in model performance with increasing computati

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

Visualizing LLM Latent Space Geometry Through Dimensionality Reduction

DGX agent

arXiv:2511.21594v3 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art results across many natural language tasks, but their internal mechanisms remain difficult to

model-releasesarxiv-cs-lg
11 Jun 2026
Model Releases

When Generic Prompt Improvements Hurt: Evaluation-Driven Iteration for LLM Applications

DGX agent

arXiv:2601.22025v2 Announce Type: replace-cross Abstract: Evaluating Large Language Model (LLM) applications differs from conventional software testing because outputs are probabilistic, semantically

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Achieving Cloud-Grade SLOs for Local Mixture-of-Experts Inference through CPU-GPU Hybrid Design

DGX agent

arXiv:2606.10493v1 Announce Type: cross Abstract: Local deployment of large Mixture-of-Experts (MoE) models falls short of the service quality achieved in cloud-scale environments, even under low-conc

model-releasesarxiv-cs-ai
10 Jun 2026
← Previous
1…391392393394395…1326
Next →