AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

PragMatch: Separating Pragmatic Incongruity from Cross-Modal Mismatch in Large Vision-Language Models

DGX agent

arXiv:2608.09772v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated strong performance on multimodal benchmarks, yet it remains unclear whether they genuinely reason

model-releasesarxiv-cs-cl
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Research

Scaling Audio Models Efficiently: A Joint Study of Compute Constraints and Optimization Behavior

DGX agent

arXiv:2606.22790v2 Announce Type: replace-cross Abstract: In this paper, we investigate the tradeoffs between compute allocation and model performance for two speech processing tasks: Automatic Speech

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

DGX agent

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries,

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

The Evolution of Mixture-of-Experts Architectures in Large Language Models: Routing, Topology, Load Balancing, and Expert Parallelism

DGX agent

arXiv:2608.08650v1 Announce Type: new Abstract: Mixture-of-Experts models increase parameter capacity while keeping the computation activated by each token bounded, but their architectural evolution c

model-releasesarxiv-cs-cl
11 Aug 2026
Research

TokenPrint: A Calibrated Token-Space Fingerprint for Language-Model Provenance

DGX agent

arXiv:2608.08139v1 Announce Type: new Abstract: Establishing the provenance of a language model---including its base checkpoint and possible overlap in training distributions---is a governance challen

researcharxiv-cs-ai
11 Aug 2026
Model Releases

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

DGX agent

arXiv:2608.07762v1 Announce Type: new Abstract: LLM benchmarks can build an organization's reputation and attract customers, but only when results are transparent and verifiable. Unverified claims tha

model-releasesarxiv-cs-ai
11 Aug 2026
Research

World Simulator: Queer Erotica and the Absurdity of AI Video Models That Promise the World

DGX agent

arXiv:2608.07510v1 Announce Type: cross Abstract: Increasingly, AI video models are marketed as 'world simulators,' suggesting their ability to model infinite realities. Despite such claims, these mod

researcharxiv-cs-cv
11 Aug 2026
Research

ZOMP: Zeroth-Order Multi-Modal Prompt Tuning for Vision-Language Models

DGX agent

arXiv:2608.08060v1 Announce Type: new Abstract: Fine-tuning vision-language models such as CLIP typically requires backpropagation (BP) through the full model, which is infeasible when only forward-pa

researcharxiv-cs-cv
11 Aug 2026
Model Releases

A Picture is Worth a Thousand Tokens: How Vision Language Models Cut AI Energy Costs While Improving Accuracy

DGX agent

arXiv:2608.07427v1 Announce Type: new Abstract: LLM inference accounts for over 90% of AI operational energy, scaling directly with input token count---a critical inefficiency for telecom network anal

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

ArchEGraph: A Large-Scale Graph Dataset for Geometry-Topology-Physics Aligned Building Energy Modeling

DGX agent

arXiv:2608.06772v1 Announce Type: new Abstract: Accurate estimation of building energy use is essential for achieving carbon neutral and sustainable buildings. To better understand the influence of de

model-releasesarxiv-cs-lg
10 Aug 2026
Safety

Do 3D Medical Foundation Models See Through MRI Artifacts? A Controlled Study of Representation Robustness

DGX agent

arXiv:2608.06613v1 Announce Type: cross Abstract: Self-supervised 3D medical foundation models are increasingly used as general-purpose feature extractors, yet their sensitivity to MRI artifacts remai

safetyarxiv-cs-ai
10 Aug 2026
Research

Flow-Corrected Shape Optimization: Taming Manifold Drift in High-Dimensional 3D Models

DGX agent

arXiv:2608.07199v1 Announce Type: new Abstract: Optimizing 3D shapes within the latent spaces of deep generative models is fundamental to computer assisted engineering, yet remains prone to a critical

researcharxiv-cs-cv
10 Aug 2026
Research

Gaze Behavior in Visual World Experiments Can be Modeled With Off-the-shelf Language-Vision Encoders

DGX agent

arXiv:2608.07282v1 Announce Type: new Abstract: The recent advances in neural language models have also spurred much work in computational psycholinguistics, asking whether neural LMs are also promisi

researcharxiv-cs-cl
10 Aug 2026
Safety

Is Forward Prediction Enough? Physical State Grounding for JEPA World Models

DGX agent

arXiv:2608.06799v1 Announce Type: cross Abstract: Learning structured and control-relevant latent representations remains a key challenge for world models. Recent JEPA-based world models learn action-

safetyarxiv-cs-cv
10 Aug 2026
Model Releases

Multi Codec Discrete Diffusion Model for Text Guided Speech Inpainting and Editing

DGX agent

arXiv:2608.06424v1 Announce Type: cross Abstract: Speech recordings often contain missing, corrupted, or incorrect regions that must be reconstructed or modified without re-synthesizing the entire utt

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

PHASE-Tree: Modeling Character-State Evolution in Long-Horizon Role-Playing Dialogue

DGX agent

arXiv:2608.06975v1 Announce Type: cross Abstract: Long-horizon role-playing demands that characters remain recognizable as they evolve with the narrative. Yet existing work falls short on two fronts:

model-releasesarxiv-cs-ai
10 Aug 2026
Tutorials

UniJEPA: A Unified Joint-Embedding Predictive Architecture for Task-Agnostic Visual World Modeling

DGX agent

arXiv:2608.07409v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) have emerged as a principled framework for self-supervised learning of world models in compact latent s

tutorialsarxiv-cs-cv
10 Aug 2026
Local Ai

ZIPBrain: Can EEG Foundation Models Be Faster, Locally Deployable, but Accurate?

DGX agent

arXiv:2608.07033v1 Announce Type: new Abstract: This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and locally deployable without sacrificing accura

local-aiarxiv-cs-ai
10 Aug 2026
Model Releases

DASH: Divergence-Adaptive Supervision Horizons for On-Policy Self-Distillation of Reasoning Models

DGX agent

arXiv:2608.06243v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves the reasoning capabilities of large language models using automatically verifiable outcom

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

DynaPix: Can Vision-Language Models Identify the Exact Future?

DGX agent

arXiv:2608.05505v1 Announce Type: new Abstract: Acting in a physical scene requires knowing its real later state, not a plausible one. Current evaluations often accept words or a realistic-looking ima

model-releasesarxiv-cs-cv
7 Aug 2026
Research

Evaluating Machine Learning Models for Post-Wildfire Debris-Flow Prediction

DGX agent

arXiv:2608.05265v1 Announce Type: cross Abstract: Prediction of post-wildfire debris flows is critical for mitigating hazards to communities, infrastructure, and resources during intense rainfall in r

researcharxiv-cs-cv
7 Aug 2026
Safety

GeniWorld: A Generalizable Interactive World Model for Robotic Manipulation via Visual Actions

DGX agent

arXiv:2608.06332v1 Announce Type: new Abstract: Generalist robot policies exhibit strong capabilities, but their robustness in complex and unseen environments remains limited. Scaling robot learning a

safetyarxiv-cs-ro
7 Aug 2026
Model Releases

Plausible Patients, Impossible Populations: Auditing Epidemiological Fidelity in Large Language Model Mental Health Simulations

DGX agent

arXiv:2604.17359v2 Announce Type: replace-cross Abstract: Language models asked to simulate psychiatric patients produce cases that survive inspection one at a time and populations that match no real

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

CLIP-CC-Bench: Evaluating Paragraph-Level Video Descriptions in Video-Language Models

DGX agent

arXiv:2608.04302v1 Announce Type: new Abstract: Benchmarking video-language models has largely focused on short clips and single-sentence metrics, leaving open whether current systems can generate acc

safetyarxiv-cs-cv
6 Aug 2026
Model Releases

Faster-WAM: Efficient Inference-Time Future Conditioning for Robust World Action Models

DGX agent

arXiv:2608.04404v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot manipulation by learning how the environment evolves beyond the current observation. However, existing approach

model-releasesarxiv-cs-cv
6 Aug 2026
Model Releases

From Financial Sentiment Classification to Return Predictability: A QLoRA Benchmark of Large Language Models

DGX agent

arXiv:2608.04200v1 Announce Type: cross Abstract: Financial sentiment classifiers are commonly evaluated against human labels, but strong linguistic performance does not necessarily imply economically

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

HyPASE: Hyperbolic Geometry for Parameter-Efficient Speech Emotion Fine-Tuning Framework for Large Audio-Language Models

DGX agent

arXiv:2608.04351v1 Announce Type: cross Abstract: Large Audio-Language Models (LALMs) excel at general speech understanding; however, adapting them to fine-grained tasks like Speech Emotion Recognitio

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

LiNC: Lightweight Noise Correction via Per-Sample Trust and Gaussian Mixture Modeling

DGX agent

arXiv:2608.04147v1 Announce Type: cross Abstract: Label noise is common in medical imaging datasets due to factors such as inter-rater variability, annotation errors, and ambiguous cases. This can sev

model-releasesarxiv-cs-ai
6 Aug 2026
Research

Mamba with Hierarchical Memory: Solving Representation Bottleneck in Long Sequence Modeling

DGX agent

arXiv:2608.02347v2 Announce Type: replace Abstract: Recurrent linear attention models (RLAs) such as Mamba offer efficient linear-time sequence modeling as an alternative to Transformers, yet their fi

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Same Formulas, Different Semantics: Do Language Models Follow Modal Logic Specifications?

DGX agent

arXiv:2608.05097v1 Announce Type: new Abstract: Reasoning about necessity and possibility depends on assumptions about accessibility between worlds and about which objects exist at each one. The same

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Teaching Foundation Models to Read mmWave: Pose-Guided Kinematic Representation for Human Behavior Understanding

DGX agent

arXiv:2608.04127v1 Announce Type: new Abstract: Large language model agents need to perceive human behavior in physical environments. Millimeter-wave (mmWave) radar provides a privacy-friendly and con

model-releasesarxiv-cs-cv
6 Aug 2026
Safety

A Blind Spot in Alignment: Quantifying Biosecurity Risks in Large Language Models

DGX agent

arXiv:2608.02684v1 Announce Type: cross Abstract: Large Language Models (LLMs) are accelerating biological research, yet this same capability poses a critical biosecurity threat: models that assist in

safetyarxiv-cs-ai
5 Aug 2026
Safety

A game theory for foundation models shows new paths to rational cooperation through similarity inference

DGX agent

arXiv:2608.03958v1 Announce Type: new Abstract: As autonomous agents powered by foundation models are increasingly integrated into social and economic systems, understanding the principles governing t

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

Can Text-to-Image Models Draw from the Right Frame of Reference?

DGX agent

arXiv:2608.03357v1 Announce Type: new Abstract: Spatial instruction following has become a crucial requirement for text-to-image (T2I) generation. A common challenge arises when directional expression

model-releasesarxiv-cs-cv
5 Aug 2026
Research

Caved or Convinced: Temporal Sampling Gates Claim Deference in Video Large Language Models

DGX agent

arXiv:2608.03160v1 Announce Type: cross Abstract: When asked which of two events came first, video large language models can fail in two opposite ways: cave to a false claim, or reject a true one. Pri

researcharxiv-cs-cv
5 Aug 2026
Model Releases

Cross-Lingual Bias in Large Language Models: A Comparative Analysis of English and Swahili

DGX agent

arXiv:2608.03532v1 Announce Type: new Abstract: Large language models are increasingly deployed in multilingual contexts, yet safety alignment and bias evaluation remain overwhelmingly English-centric

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

dots.tts.edit: Precisely Controlled Speech Editing with a Continuous Autoregressive Model

DGX agent

arXiv:2608.02673v1 Announce Type: cross Abstract: Speech editing for content creation requires precise control over both what an edit should do and where it should apply. Free-form natural language pr

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

DGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

model-releasesarxiv-cs-cl
5 Aug 2026
Hardware

LiLa-WAM: Lightweight Latent Reasoning World-Action Model for Robotic Manipulation

DGX agent

arXiv:2608.03701v1 Announce Type: cross Abstract: World-action modeling has emerged as a promising paradigm for robotic control, as it empowers models to go beyond reacting to observations and anticip

hardwarearxiv-cs-ai
5 Aug 2026
Research

MDLMPE: Distribution Aware Positional Encoding for Masked Diffusion Language Models

DGX agent

arXiv:2608.03769v1 Announce Type: cross Abstract: Masked diffusion language models (MDLMs) enable parallel generation and bidirectional context modeling, but their positional context differs fundament

researcharxiv-cs-ai
5 Aug 2026
Model Releases

MIDI-LLM: Improving Text-to-MIDI Music Generation via Adapting Large Language Models

DGX agent

arXiv:2511.03942v2 Announce Type: replace-cross Abstract: We present MIDI-LLM, a recipe that improves multitrack text-to-MIDI generation via adapting Large Language Models (LLMs). MIDI-LLM expands an

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Qwen-3D: A Generalist 3D Vision-Language Model for Spatial Understanding

DGX agent

arXiv:2608.02980v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have achieved remarkable success on images and short videos, yet scaling them to long videos remains challenging due to f

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

Self-Guided Adaptive Safety Alignment: Synthesizing and Internalizing Guidelines in Reasoning Models

DGX agent

arXiv:2511.21214v4 Announce Type: replace-cross Abstract: Explicit safety policies can improve reasoning-model safety, but their effective coverage may lag behind evolving jailbreak strategies. We stu

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

SeqLLM: Augmenting LLMs with Behavioral-Sequence Modeling for High-Stakes Decisions at WeChat Pay

DGX agent

arXiv:2608.03063v1 Announce Type: new Abstract: Merchant risk control at large payment platforms screens tens of millions of merchants daily, where false positives harm legitimate merchants and false

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

SlimVLM: Sensitivity-aware Dynamic Structured Pruning with Adaptive Visual Token Selection for Efficient Vision-Language Models

DGX agent

arXiv:2608.03580v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have demonstrated remarkable performance in processing and understanding both text and images, their large parameter

model-releasesarxiv-cs-cv
5 Aug 2026
Safety

SP3O: Reinforcement Learning from Segment Preferences without Reward Modeling

DGX agent

arXiv:2608.02951v1 Announce Type: cross Abstract: Preference-based reinforcement learning (PbRL) for general stochastic MDPs often requires training a reward model. Existing reward-model-free methods

safetyarxiv-cs-ai
5 Aug 2026
Safety

A Heuristic Perspective on Debiasing Language Models

DGX agent

arXiv:2608.00622v1 Announce Type: new Abstract: Language models (LMs) often acquire various biases during pre-training and may express them in interactions, potentially causing social harm. Existing m

safetyarxiv-cs-cl
4 Aug 2026
← Previous
1…7677787980…1030
Next →