AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Research

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling

DGX agent

arXiv:2603.22911v2 Announce Type: replace-cross Abstract: Due to the great saving of computation and memory overhead, token compression has become a research hot-spot for MLLMs and achieved remarkable

researcharxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models

DGX agent

arXiv:2601.03416v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have become widely deployed, yet their safety alignment remains fragile under adversarial inputs. Previous

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

MLLM-as-a-Judge Exhibits Model Preference Bias

DGX agent

arXiv:2604.11589v1 Announce Type: new Abstract: Automatic evaluation using multimodal large language models (MLLMs), commonly referred to as MLLM-as-a-Judge, has been widely used to measure model perf

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling

DGX agent

arXiv:2604.09580v1 Announce Type: new Abstract: Standard Chain-of-Thought (CoT) prompting empowers Large Language Models (LLMs) with reasoning capabilities, yet its reliance on linear natural language

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Quantitative Introspection in Language Models: Tracking Emotive States Across Conversation

DGX agent

arXiv:2603.18893v2 Announce Type: replace Abstract: Tracking the internal states of large language models across conversations is important for safety, interpretability, and model welfare, yet current

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Seeing No Evil: Blinding Large Vision-Language Models to Safety Instructions via Adversarial Attention Hijacking

DGX agent

arXiv:2604.10299v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) rely on attention-based retrieval of safety instructions to maintain alignment during generation. Existing attack

model-releasesarxiv-cs-cl
14 Apr 2026
Research

Suiren-1.0 Technical Report: A Family of Molecular Foundation Models

DGX agent

arXiv:2603.21942v2 Announce Type: replace-cross Abstract: We introduce Suiren-1.0, a family of molecular foundation models for the accurate modeling of diverse organic systems. Suiren-1.0 comprising t

researcharxiv-cs-ai
14 Apr 2026
Model Releases

TS-Haystack: A Multi-Scale Retrieval Benchmark for Time Series Language Models

DGX agent

arXiv:2602.14200v4 Announce Type: replace Abstract: Time Series Language Models (TSLMs) are emerging as unified models for reasoning over continuous signals in natural language. However, long-context

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

Woosh: A Sound Effects Foundation Model

DGX agent

arXiv:2604.01929v2 Announce Type: replace-cross Abstract: The audio research community depends on open generative models as foundational tools for building novel approaches and establishing baselines.

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Advantage-Guided Diffusion for Model-Based Reinforcement Learning

DGX agent

arXiv:2604.09035v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL) with autoregressive world models suffers from compounding errors, whereas diffusion world models mitigate this

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

AgriChain Visually Grounded Expert Verified Reasoning for Interpretable Agricultural Vision Language Models

DGX agent

arXiv:2604.07814v1 Announce Type: new Abstract: Accurate and interpretable plant disease diagnosis remains a major challenge for vision-language models (VLMs) in real-world agriculture. We introduce A

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

An empirical study of LoRA-based fine-tuning of large language models for automated test case generation

DGX agent

arXiv:2604.06946v1 Announce Type: cross Abstract: Automated test case generation from natural language requirements remains a challenging problem in software engineering due to the ambiguity of requir

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

DGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Multi-objective Evolutionary Merging Enables Efficient Reasoning Models

DGX agent

arXiv:2604.06465v1 Announce Type: cross Abstract: Reasoning models have demonstrated remarkable capabilities in solving complex problems by leveraging long chains of thought. However, this more delibe

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

OmniTabBench: Mapping the Empirical Frontiers of GBDTs, Neural Networks, and Foundation Models for Tabular Data at Scale

DGX agent

arXiv:2604.06814v1 Announce Type: cross Abstract: While traditional tree-based ensemble methods have long dominated tabular tasks, deep neural networks and emerging foundation models have challenged t

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

oslash Source Models Leak What They Shouldn't nrightarrow: Unlearning Zero-Shot Transfer in Domain Adaptation Through Adversarial Optimization

DGX agent

arXiv:2604.08238v1 Announce Type: new Abstract: The increasing adaptation of vision models across domains, such as satellite imagery and medical scans, has raised an emerging privacy risk: models may

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Small Vision-Language Models are Smart Compressors for Long Video Understanding

DGX agent

arXiv:2604.08120v1 Announce Type: cross Abstract: Adapting Multimodal Large Language Models (MLLMs) for hour-long videos is bottlenecked by context limits. Dense visual streams saturate token budgets

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

DGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

DGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

model-releasesarxiv-cs-ai
10 Apr 2026
Research

WorldMAP: Bootstrapping Vision-Language Navigation Trajectory Prediction with Generative World Models

DGX agent

arXiv:2604.07957v1 Announce Type: cross Abstract: Vision-language models (VLMs) and generative world models are opening new opportunities for embodied navigation. VLMs are increasingly used as direct

researcharxiv-cs-cv
10 Apr 2026
Model Releases

Flex-pi: A Multi-Stream World-Action Model with Compute Flexibility

DGX agent

arXiv:2608.10860v1 Announce Type: cross Abstract: World-action models (WAMs) predict the future to act better, but nearly all of them predict only RGB latents, trained purely for pixel reconstruction,

model-releasesarxiv-cs-cv
12 Aug 2026
Agents

Improving Constraint Models with LLM Agents

DGX agent

arXiv:2608.08127v1 Announce Type: new Abstract: The runtime of Constraint Programming (CP) solvers is highly sensitive to modeling choices, such as symmetry breaking, implied constraints, global const

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Learning How the World Evolves: Extrapolative Video World Models via Latent Dynamics Reasoning

DGX agent

arXiv:2608.09926v1 Announce Type: new Abstract: The world evolves following its dynamics, i.e., its laws of motion. However, leading video diffusion models largely fit the pixels without modeling how

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

DGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

model-releasesarxiv-cs-ai
11 Aug 2026
Applications

AtlasVLA: Persistent World-Ego State Modeling for Vision-Language-Action Models

DGX agent

arXiv:2608.06729v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models have advanced embodied AI, their fundamentally reactive paradigm severely limits performance in partially ob

applicationsarxiv-cs-cv
10 Aug 2026
Model Releases

Capacity Confounds and Coverage Guarantees in Adaptive Sub-model Federated Learning

DGX agent

arXiv:2608.07157v1 Announce Type: new Abstract: Sub-model federated learning lets resource-constrained clients train width-reduced versions of a global model, but existing methods allocate capacity by

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

Capek 0.5: An Execution-Centric Vision-Language Model for Embodied Intelligence

DGX agent

arXiv:2608.06756v1 Announce Type: new Abstract: Vision-language models are increasingly serving as the reasoning core of embodied agents. Robot execution is inherently iterative: each action reshapes

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Divergent Response Modes in Frontier Language Models Under Steering Pressure

DGX agent

arXiv:2608.06578v1 Announce Type: new Abstract: Frontier language models are trained using distinct data, objectives, and safety pipelines. Whether these differences produce measurably different behav

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Newton-Schulz Retraction-Based Inference Enables Hidden Quantum Markov Models to Outperform Classical HMMs

DGX agent

arXiv:2608.06554v1 Announce Type: new Abstract: Hidden Markov models (HMMs) are widely used probabilistic models for discrete sequential data but can be limited when hidden dynamics are complex. Hidde

model-releasesarxiv-cs-lg
10 Aug 2026
Research

RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction

DGX agent

arXiv:2608.06310v1 Announce Type: cross Abstract: Recent advances in reward modeling show a paradigm shift from discriminative reward models to generative reward models. However, despite their strong

researcharxiv-cs-cl
7 Aug 2026
Model Releases

BnBERT-iPET: Sparse Few-Shot Language Modeling for Bengali via Lottery Ticket Pruning

DGX agent

arXiv:2608.05104v1 Announce Type: new Abstract: Deep neural networks have shown impressive success in NLP tasks owing to their complex structure and huge number of edges. Achieving state-of-the-art pe

model-releasesarxiv-cs-lg
6 Aug 2026
Research

Generative Models for Modeling and Synthesizing MIMO Channels in Adverse Weather Conditions

DGX agent

arXiv:2608.00156v1 Announce Type: cross Abstract: The push for broader coverage in future cellular networks depends on reliable service, yet this is increasingly harder to do as we encounter more inst

researcharxiv-cs-lg
4 Aug 2026
Model Releases

Prompt-Induced Waste in Large Reasoning Models: A Preregistered Two-Harness Benchmark of Coding Agents

DGX agent

arXiv:2608.01347v1 Announce Type: new Abstract: Large reasoning models used as coding agents incur costs from deliberation, tool calls, and repeated agent turns, yet the causal effect of prompt wordin

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Recursive Vision Language Models for General Symbolic Reasoning

DGX agent

arXiv:2608.01534v1 Announce Type: new Abstract: Hard symbolic-reasoning tasks such as Sudoku, maze pathfinding, and ARC remain challenging for LLMs due to their fixed-depth autoregressive reasoning, w

model-releasesarxiv-cs-cv
4 Aug 2026
Research

Sixteen models, fewer than two voices: measuring ensemble dispersion where no answer is uniquely correct

DGX agent

arXiv:2608.00285v1 Announce Type: new Abstract: Sixteen language models drawn from ten families produced, on average, the semantic diversity of 1.69 distinct formulations of a psychotherapeutic case,

researcharxiv-cs-cl
4 Aug 2026
Model Releases

TextNCA: Neural Cellular Automata for Language Modeling via Hierarchical Local Attention

DGX agent

arXiv:2608.02050v1 Announce Type: new Abstract: Can a strictly local, iterated, weight-shared computation primitive support language modelling, and which of those three properties actually drives the

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

What Carries the Signal in Pathology Foundation-Model Atlases? A Patient-Level Controlled Benchmark in Breast Cancer

DGX agent

arXiv:2608.00105v1 Announce Type: new Abstract: Pathology foundation models are reported to encode molecular programmes in tissue morphology, but the evidence is usually a cohort-wide ranked gene list

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

PARALLEL: A Prefrontal-Aligned Reinforcement inspired Approach for Language-Model Learning under Explicit Limits

DGX agent

arXiv:2607.28982v1 Announce Type: cross Abstract: Recent language models achieve strong performance across a variety of tasks, but conventional adaptation applies updates uniformly across training sam

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

SAM+D: Parameter-Efficient Dimensional Lifting of SAM-Family Models via Depth-Routed LoRA and Depth Shifting

DGX agent

arXiv:2607.29033v1 Announce Type: new Abstract: Existing methods for adapting 2D foundation models such as SAM to 3D volumes either process slices independently---ignoring inter-slice context---or req

model-releasesarxiv-cs-cv
3 Aug 2026
Research

Shall We Play a Game? Language Models for Open-ended Wargames

DGX agent

arXiv:2509.17192v3 Announce Type: replace Abstract: LLM-based social simulations can make a generated transcript look like a single behavioral signal, but the model behind that transcript may be doing

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Comparison of a Parametric Physics-Informed Neural Network and a Tensorial Reduced-Order Model for the Shallow-Water Dam-Break Problem

DGX agent

arXiv:2607.27433v1 Announce Type: cross Abstract: We develop two parametric data-driven reduced models: a physics-informed neural network (PINN) and a non-intrusive tensorial reduced-order model (TROM

model-releasesarxiv-cs-lg
31 Jul 2026
Tutorials

PhiZero: A World Model Built Around Physical Language

DGX agent

arXiv:2607.28624v1 Announce Type: new Abstract: We introduce PhiZero, a physical world model built around physical language, a compact discrete representation of world-state transitions. Existing phys

tutorialsarxiv-cs-cv
31 Jul 2026
Model Releases

Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning

DGX agent

arXiv:2607.28478v1 Announce Type: new Abstract: As large language models (LLMs) continue to advance in complex reasoning tasks, they have learned to heavily prioritize explicit conditions provided in

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

BayesAME: Bayesian Active Model Evaluation

DGX agent

arXiv:2607.27023v1 Announce Type: new Abstract: Evaluating large generative models across benchmarks is time-consuming and computationally expensive. This drives the need for methods that can estimate

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

ForgetBench: Benchmarking Forgetting Dynamics of Long-Term Parametric Memory in Language Models

DGX agent

arXiv:2607.26455v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong capabilities in knowledge acquisition and reasoning, yet their ability to retain previously acquir

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text

DGX agent

arXiv:2607.26751v1 Announce Type: new Abstract: State-of-the-art intracortical brain-to-text systems pair a neural-sequence phone decoder with an external language model. Two design axes remain undere

model-releasesarxiv-cs-cl
30 Jul 2026
Tutorials

StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation

DGX agent

arXiv:2607.26754v1 Announce Type: new Abstract: Recent game world models can generate visually realistic and interactive environments conditioned on player actions. However, games are not defined by p

tutorialsarxiv-cs-cv
30 Jul 2026
← Previous
1…1516171819…1012
Next →