AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
3 Jun 2026

Patcher: Post-Hoc Patching of Backdoored Large Language Models

Local AiDGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

Pretraining Language Models on Historical Text

Model ReleasesDGX agent

arXiv:2606.02991v1 Announce Type: cross Abstract: We introduce TypewriterLM, a 7.24B History language model (LM) trained exclusively on English text predating 1913. Developing History LMs requires add

TurtleAI: Benchmarking Multimodal Models for Visual Programming in Turtle Graphics

Model ReleasesDGX agent

arXiv:2606.03626v1 Announce Type: cross Abstract: Vision-language models (VLMs) have been explored for visual programming, where they generate code to solve visual tasks. However, most prior work focu

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours.

TutorialsDGX agent

Working with @FireworksAI_HQ to make MAI models easy to fine-tune and fully yours. Microsoft MAI models. Coming soon to Fireworks. Intelligence you control. End-to-end lineage you can prove. Fine-tune

2 Jun 2026

A Foundation Model for Wearable Movement Data in Mental Health Research

ApplicationsDGX agent

arXiv:2411.15240v5 Announce Type: replace-cross Abstract: Wearable movement data is collected by nearly all commercially available smartwatches and is a valuable resource for mental health research, r

Accuracy, Stability, and Repeated-Run Reliability of Large Language Models on Deterministic Programming Tasks

Model ReleasesDGX agent

arXiv:2606.00920v1 Announce Type: cross Abstract: Run-level pass rate overstates retry-free coverage by up to 17.8 percentage points -- and the gap is largest precisely for mid-performing systems. We

Active Exploring like a Pigeon: Reinforcing Spatial Reasoning via Agentic Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.02459v1 Announce Type: new Abstract: Enabling Vision-Language Models (VLMs) to perform spatial reasoning remains challenging. Existing approaches treat VLMs as passive observers, which is d

Benchmarking Waitlist Mortality Prediction in Heart Transplantation Through Time-to-Event Modeling using New Longitudinal UNOS Dataset

Model ReleasesDGX agent

arXiv:2507.07339v2 Announce Type: replace-cross Abstract: Decisions about managing patients on the heart transplant waitlist are currently made by committees of doctors who consider multiple factors,

CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection

Model ReleasesDGX agent

arXiv:2606.00101v1 Announce Type: cross Abstract: With the rapid advancement of artificial intelligence generated content (AIGC) technologies, video forgery has become increasingly prevalent, posing n

COMAP: Co-Evolving World Models and Agent Policies for LLM Agents

SafetyDGX agent

arXiv:2606.02372v1 Announce Type: new Abstract: Equipping language agents with world models enables them to anticipate environment dynamics and evaluate candidate actions before execution. However, ex

EST-PRM: Stress-Testing Process Reward Models Before They Become Load-Bearing

ResearchDGX agent

arXiv:2606.00437v1 Announce Type: new Abstract: Process reward models (PRMs) are widely used in language-model training with dense step-level supervision. They assume PRM scores are stable proxies for

FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning

SafetyDGX agent

arXiv:2510.09222v3 Announce Type: replace Abstract: Flow Matching (FM) has shown remarkable ability in modeling complex distributions and achieves strong performance in offline imitation learning for

From question to model. The public equity investing plugin for Codex.

Model ReleasesDGX agent

This post describes OpenAI's public equity investing plugin for Codex, which enables users to convert investment questions into analytical models through natural language processing. The plugin likely

From Zero to Hero: Training-Free Custom Concept Spawning in World Models

ResearchDGX agent

arXiv:2606.02575v1 Announce Type: new Abstract: Autoregressive world models have emerged as a powerful paradigm for interactive video generation, allowing users to navigate dynamically generated envir

Hybrid Neural Ordinary Differential Equations for Data-Efficient Polymerization Modeling with Incomplete Kinetics

ApplicationsDGX agent

arXiv:2606.02145v1 Announce Type: new Abstract: Accurate prediction of polymerization dynamics is essential for process design, control, and optimization. Yet, purely mechanistic models require labor-

LASER: Loss-Aware Singular-value Decomposition and Rank Allocation for Efficient Low-Precision Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.00573v1 Announce Type: new Abstract: Vision-language models (VLMs) deliver strong multimodal reasoning capabilities, but their large computational cost and high parameter counts make deploy

Learning to Remember, Learn, and Forget in Attention-Based Models

Model ReleasesDGX agent

arXiv:2602.09075v3 Announce Type: replace-cross Abstract: In-Context Learning (ICL) in transformers acts as an online associative memory and is believed to underpin their high performance on complex s

One Bias After Another: Mechanistic Reward Shaping and Persistent Biases in Language Reward Models

SafetyDGX agent

arXiv:2603.03291v2 Announce Type: replace-cross Abstract: Reward Models (RMs) are crucial for online alignment of language models (LMs) with human preferences. However, RM-based preference-tuning is v

Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models

Model ReleasesDGX agent

arXiv:2606.00284v1 Announce Type: new Abstract: While continual pretraining~(CPT) is a practical way to extend large language models to new languages, naive finetuning on targeted data erodes existing

Physical Object Understanding with a Physically Controllable World Model

ResearchDGX agent

arXiv:2606.00439v1 Announce Type: new Abstract: A central challenge in visual intelligence is learning the physical structure of scenes from raw videos: how regions form objects and the laws that gove

Position: Good Embodied Reward Models Need Bad Behavior Data

SafetyDGX agent

arXiv:2606.01036v1 Announce Type: new Abstract: This position paper argues that to obtain reliable embodied reward models, the community must invest in ``bad'' robot data: failed, suboptimal, error-pr

ProactiveLLM: Learning Active Interaction for Streaming Large Language Models

Local AiDGX agent

arXiv:2606.00523v1 Announce Type: new Abstract: Standard Large Language Models (LLMs) follow a read-then-generate paradigm, causing unnecessary latency and computation. Streaming LLMs alleviate this i

Query-Limited Community Recovery in Stochastic Block Models

Model ReleasesDGX agent

arXiv:2606.02055v1 Announce Type: cross Abstract: We study exact community recovery in the two-community stochastic block model on n vertices under limited and noisy access to network data. The learne

RealityTest: How People Probe AI Identity and Whether Models Disclose It

Model ReleasesDGX agent

arXiv:2606.00168v1 Announce Type: new Abstract: AI systems are increasingly deployed in conversational settings where users may be uncertain whether they are speaking with a human or an AI. Despite mo

RefLoRA: Refactored Low-Rank Adaptation for Efficient Fine-Tuning of Large Models

Model ReleasesDGX agent

arXiv:2505.18877v4 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) lowers the computational and memory overhead of fine-tuning large models by updating a low-dimensional subspace of the pr

Score-Control for Hallucination Reduction in Diffusion Models

Model ReleasesDGX agent

arXiv:2606.00377v1 Announce Type: new Abstract: Diffusion models have emerged as the backbone of modern generative AI, powering advances in vision, language, audio and other modalities. Despite their

Strategizing at Speed: A Learned Model Predictive Game for Multi-Agent Drone Racing

Model ReleasesDGX agent

arXiv:2602.06925v2 Announce Type: replace Abstract: Autonomous drone racing pushes the boundaries of high-speed motion planning and multi-agent strategic decision-making. Success in this domain requir

Task Structure Reverses Layerwise State Encoding in Sequence Models

ResearchDGX agent

arXiv:2606.00926v1 Announce Type: cross Abstract: Mechanistic studies of sequence models often treat layerwise state encodings as architectural traits: recurrent models concentrate readable state, att

Towards a Physics Foundation Model

TutorialsDGX agent

arXiv:2509.13805v4 Announce Type: replace-cross Abstract: Foundation models have revolutionized natural language processing through a ``train once, deploy anywhere'' paradigm, where a single pre-train

TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models

Model ReleasesDGX agent

arXiv:2606.00023v1 Announce Type: cross Abstract: The rapid development of Language Diffusion Models (LDMs) challenges the dominant position of auto-regressive competitors in language processing. Howe

WAON: A Large-Scale Japanese Image-Text Dataset for Cultural Adaptation in Contrastive Vision-Language Models

Model ReleasesDGX agent

arXiv:2510.22276v3 Announce Type: replace-cross Abstract: Contrastive vision-language models have achieved remarkable progress through large-scale pretraining. Recent work has shown that removing Engl

When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

Model ReleasesDGX agent

arXiv:2602.08236v2 Announce Type: replace-cross Abstract: Despite rapid progress in MLLMs, visual spatial reasoning remains unreliable when correct answers depend on how a scene would appear under uns

World Models for Robotic Manipulation: A Survey

SafetyDGX agent

arXiv:2606.00113v1 Announce Type: new Abstract: Robotic manipulation depends on the ability to anticipate how actions reshape objects, contacts, and scene geometry before execution. Learned world mode

X-Foresight: A Joint Vision-Action Causal Forecasting Network via Predictive World Modeling

SafetyDGX agent

arXiv:2605.24892v2 Announce Type: replace Abstract: Physical world knowledge resides mainly in videos. Equipping Vision-Language-Action (VLA) models with such knowledge is fundamental for safe and gen

1 Jun 2026

BenHalluEval: A Multi-Task Hallucination Evaluation Framework for Large Language Models on Bengali

Model ReleasesDGX agent

arXiv:2605.31483v1 Announce Type: new Abstract: Despite Bengali being the sixth most spoken language in the world, no prior work has systematically evaluated hallucination in large language models (LL

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation

Model ReleasesDGX agent

arXiv:2605.31286v1 Announce Type: cross Abstract: Real-world household robots require Vision-Language-Action (VLA) foundation models that can acquire reusable manipulation skills across diverse object

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion

Model ReleasesDGX agent

arXiv:2605.31170v1 Announce Type: cross Abstract: Monitoring autonomous language model agents currently relies mostly on surface behavior. But what happens when agent populations invent new languages

End-to-End Compression for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2602.05649v2 Announce Type: replace Abstract: The long-standing dominance of gradient-boosted decision trees for tabular data has recently been challenged by in-context learning tabular foundati

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video…

HardwareDGX agent

🆕Grok Imagine’s Video Agent Moment: Cosmos, xAI, World Models, Generative UI, & the Codex Phase for Video! https://www.latent.space/p/video-agents @EthanHe_42, former @xai world model lead and @nvidia

Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback

Model ReleasesDGX agent

arXiv:2605.30478v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) trains language models using programmatically checkable signals such as unit-test outcomes, enab

Motion Tracking with Muscles: Predictive Control of a Parametric Musculoskeletal Canine Model

ResearchDGX agent

arXiv:2506.23768v2 Announce Type: replace Abstract: We introduce a novel musculoskeletal model of a dog, procedurally generated from accurate 3D muscle meshes. Accompanying this model is a motion capt

PatchWorld: Gradient-Free Optimization of Executable World Models

Local AiDGX agent

arXiv:2605.30880v1 Announce Type: cross Abstract: Text-agent environments are typically modeled as partially observable Markov decision processes (POMDPs), assuming that the simulator's latent state a

The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement

SafetyDGX agent

arXiv:2605.30888v1 Announce Type: new Abstract: Building strong reward models (RMs) for language model alignment is bottlenecked by the cost and difficulty of acquiring diverse and reliable preference

31 May 2026

Experts say ChatGPT, Gemini, and other Western AI models are turbocharging Iran's cyber operations, helping it develop malware and launch phishing attacks (Jacob Judah/Financial Times)

Model ReleasesDGX agent

Jacob Judah / Financial Times: Experts say ChatGPT, Gemini, and other Western AI models are turbocharging Iran's cyber operations, helping it develop malware and launch phishing attacks — Western AI m

29 May 2026

Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models

Model ReleasesDGX agent

arXiv:2605.30038v1 Announce Type: cross Abstract: Diffusion models generate highly realistic images but often struggle with precise text-image alignment. While recent post-training methods improve ali

AMDP: Asynchronous Multi-Directional Pipeline Parallelism for Large-Scale Models Training

Model ReleasesDGX agent

arXiv:2605.29664v1 Announce Type: cross Abstract: Pipeline parallelism is essential for large-scale model training, but existing asynchronous approaches often degrade convergence due to parameter mism

An End-to-End PyTorch Interface for Differentiable PDE Solvers: A RANS Model-Correction Study

Model ReleasesDGX agent

arXiv:2605.28858v1 Announce Type: cross Abstract: This work presents an end-to-end strategy for solving inverse problems constrained by Partial Differential Equations within a fully differentiable Mac

BioArc: Discovering Optimal Neural Architectures for Biological Foundation Models

TutorialsDGX agent

arXiv:2512.00283v3 Announce Type: replace-cross Abstract: Foundation models have revolutionized various fields such as natural language processing (NLP) and computer vision (CV). While efforts have be

Enhancing Membership Inference Attacks on Diffusion Models from a Frequency-Domain Perspective

ResearchDGX agent

arXiv:2505.20955v4 Announce Type: replace-cross Abstract: Diffusion models have achieved tremendous success in image generation, but they also raise significant concerns regarding privacy and copyrigh

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models

SafetyDGX agent

arXiv:2605.29303v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) followed by reinforcement learning (RL) has become a standard post-training paradigm for large language models. This paradi

Evaluating Dataset Watermarking for Fine-tuning Traceability of Customized Diffusion Models: A Comprehensive Benchmark and Removal Approach

Model ReleasesDGX agent

arXiv:2511.19316v2 Announce Type: replace-cross Abstract: Recent fine-tuning techniques for diffusion models enable them to reproduce specific image sets, such as particular faces or artistic styles,

Evolutionary Rule Extraction from Corporate Default Prediction Models

Model ReleasesDGX agent

arXiv:2605.29478v1 Announce Type: cross Abstract: Small and medium-sized enterprises (SMEs) represent the majority of firms in most economies and often face financial constraints and higher vulnerabil

Feature Geometry of LoRA Adapters: A Sparse Autoencoder Analysis of Representational Divergence in Fine-Tuned Language Models

Model ReleasesDGX agent

arXiv:2605.28896v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has emerged as a widely adopted approach for adapting large language models, yet the internal representational changes induce

How's it going? Reinforcement learning in language models recruits a functional welfare axis

SafetyDGX agent

arXiv:2605.30232v1 Announce Type: cross Abstract: How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation

Is Your Diffusion Sampler Actually Correct? A Sampler-Centric Evaluation of Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2602.19619v2 Announce Type: replace Abstract: Discrete diffusion language models (dLLMs) provide a fast and flexible alternative to autoregressive models (ARMs) via iterative denoising with para

Large language models reorganize representational geometry during in-context learning

Model ReleasesDGX agent

arXiv:2605.28854v1 Announce Type: new Abstract: Large language models (LLMs) exhibit remarkable flexibility: they can adapt to novel tasks from in-context examples without any parameter updates, a cap

Midpoint Generative Models

ResearchDGX agent

arXiv:2605.29920v1 Announce Type: new Abstract: We introduce Midpoint Generative Models (MGM), a principled framework for training one-step generative models. MGM is based on a simple symmetry of Flow

Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion

ResearchDGX agent

arXiv:2605.29591v1 Announce Type: new Abstract: Modeling the interplay between external stimuli and internal neural representations is a pivotal research area for Brain-Computer Interfaces (BCIs). A m

Model Fusion via Retrofitting

SafetyDGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

Multi-Turn Adaptive Prompting Attack on Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2602.14399v2 Announce Type: replace Abstract: Multi-turn jailbreak attacks have proven effective against text-only large language models (LLMs), where malicious content is gradually introduced t

← Previous
1…8586878889…1009
Next →