AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,002 results
10 Jun 2026

Easy solution to slow down recursive AI self improvement: - The lab with the top-ranked model must agree THEY must not use it for working on…

TutorialsDGX agent

Easy solution to slow down recursive AI self improvement: - The lab with the top-ranked model must agree THEY must not use it for working on frontier AI - But everyone else should have access to it. B

Exact Functional ANOVA Decomposition for Categorical Inputs Models

ResearchDGX agent

arXiv:2603.02673v2 Announce Type: replace-cross Abstract: Functional ANOVA offers a principled framework for interpretability by decomposing a model's prediction into main effects and higher-order int

Flow Control: Steering Vision-Language-Action Models with Simple Real-Time Inputs

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.10180v1 Announce Type: cross Abstract: We introduce flow control of vision-language-action (VLA) models, a simple and effective way to steer VLA actions in real-time through generic inputs,

Geometric Coastline Localization using Vision-Language Models

Local AiDGX agent

arXiv:2606.10468v1 Announce Type: new Abstract: Coastline detection in remote sensing imagery is commonly formulated as a pixel-wise segmentation problem, where the final coastline is extracted from a

HiMem-WAM: Hierarchical Memory-Gated World Action Models for Robotic Manipulation

ApplicationsDGX agent

arXiv:2606.10363v1 Announce Type: new Abstract: World Action Models (WAMs) have emerged as a new powerful paradigm for embodied intelligence, learning action-relevant visual dynamics that significantl

I really don't like the feeling of anger man. It sucks. Im really disappointed by anthropic. I've been writing model code since 2010, everyt…

TutorialsDGX agent

I really don't like the feeling of anger man. It sucks. Im really disappointed by anthropic. I've been writing model code since 2010, everything I do looks like 'frontier research'. Like genuinely. Th

Integrating Biological-Informed Recurrent Neural Networks for Glucose-Insulin Dynamics Modeling

ResearchDGX agent

arXiv:2503.19158v3 Announce Type: replace Abstract: Type 1 Diabetes (T1D) management is a complex task due to many variability factors. Artificial Pancreas (AP) systems have alleviated patient burden

Interpreting and Steering a Text-to-Speech Language Model with Sparse Autoencoders

ResearchDGX agent

arXiv:2606.10029v1 Announce Type: cross Abstract: Language models increasingly serve as the backbone of text-to-speech (TTS) systems, yet we understand little about the representations they build when

Model-Based Reinforcement Learning in Discrete-Action Non-Markovian Reward Decision Processes

ResearchDGX agent

arXiv:2512.14617v2 Announce Type: replace-cross Abstract: Many practical decision-making problems involve tasks whose success depends on the entire system history, rather than on achieving a state wit

MODIP: Efficient Model-Based Optimization for Diffusion Policies

SafetyDGX agent

arXiv:2606.10825v1 Announce Type: new Abstract: Diffusion policies (DPs) have emerged as expressive policy representations for robot learning, often used with imitation learning methods such as behavi

Predicting Future Behaviors in Reasoning Models Enables Better Steering

ResearchDGX agent

arXiv:2606.11172v1 Announce Type: new Abstract: Deployed large reasoning models (LRMs) often behave unexpectedly. Test-time steering controls LRM outputs by intervening on their hidden representations

Prefilling-dLLM: Predictive Prefilling for Long-Context Inference in Diffusion Language Models

ResearchDGX agent

arXiv:2606.10537v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) re-encode the entire prefix at every denoising step, causing recomputation that scales quadratically with contex

Sources: Trump administration officials have told CAISI to halt publication of its model assessments while an EO President Trump signed last week is implemented (Amrith Ramkumar/Wall Street Journal)

SafetyDGX agent

Amrith Ramkumar / Wall Street Journal: Sources: Trump administration officials have told CAISI to halt publication of its model assessments while an EO President Trump signed last week is implemented

Speaker Group Encoding in Self-supervised Speech Recognition Models

SafetyDGX agent

arXiv:2606.10654v1 Announce Type: new Abstract: We investigate what self-supervised speech recognition models (S3Ms) learn about speaker groups (SGs). We examine several states of S3Ms: pretrained, fi

The core problem with open weights is that the business model of frontier open weights AI does not look like open source, as there are very …

ApplicationsDGX agent

The core problem with open weights is that the business model of frontier open weights AI does not look like open source, as there are very few cases where you can make money from closed ancillary ser

The Inference Alpha: Maximizing Frontier Models on AMD

IndustryDGX agent

This article discusses strategies for optimizing the performance of advanced AI frontier models when running on AMD hardware infrastructure. It likely covers deployment best practices, hardware config

This is why frontier open models are crucial. This is extremely sad for the research community.

ResearchDGX agent

This is why frontier open models are crucial. This is extremely sad for the research community. mythos will be bad ON PURPOSE on ai 'frontier llm research' tasks, this is very very sad for the researc

Time Series as Language: A Universal Tokenizer for General-Purpose Time Series Foundation Models

ResearchDGX agent

arXiv:2606.09861v1 Announce Type: cross Abstract: While Next-Token Prediction (NTP) has unified LLM pretraining, its adaptation to unbounded, continuous time series (TS) remains open. To bridge the ga

Towards Diverse Scientific Hypothesis Search with Large Language Models

ResearchDGX agent

arXiv:2606.10587v1 Announce Type: cross Abstract: Large language models (LLMs) are on the rise for accelerating scientific discovery, most recently in advanced tasks such as generating valid scientifi

VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2606.10568v1 Announce Type: new Abstract: Vision-language-action (VLA) models have shown strong promise for robotic manipulation, but their reliability at test time remains limited by one-shot a

Vision-Assisted Foundation Model for Solving Multi-Task Vehicle Routing Problems

TutorialsDGX agent

arXiv:2606.10431v1 Announce Type: cross Abstract: Multi-task vehicle routing problems play a critical role in enhancing efficiency across various industries and service sectors. These problems consist

What is the best open sourced image model?

Local AiDGX agent

The best open-source image generation models in 2026 include FLUX.1 [schnell], Stable Diffusion 3.5 Large, HiDream-I1-Full, SANA-Sprint 1.6B, and HunyuanImage-3.0 . FLUX.1 [dev] holds the crown for ph

WMG acquires AI attribution startup Sureel, as it seeks to track when its songs and recordings are used in the training of AI models or in AI-generated works (Kristin Robinson/Billboard)

IndustryDGX agent

Kristin Robinson / Billboard: WMG acquires AI attribution startup Sureel, as it seeks to track when its songs and recordings are used in the training of AI models or in AI-generated works — The deal i

You should really all try Opus Fable low mode versus high or ultracost It's a good enough model to semi one-shot most things and this uses f…

IndustryDGX agent

You should really all try Opus Fable low mode versus high or ultracost It's a good enough model to semi one-shot most things and this uses far fewer tokens You can even tell it to figure out the most

9 Jun 2026

As frontier models (e.g. Fable 5) continue to push the task horizon of knowledge work automation, it becomes ever more important for humans …

AgentsDGX agent

As frontier models (e.g. Fable 5) continue to push the task horizon of knowledge work automation, it becomes ever more important for humans to be able to audit decisions back to the source context. It

Brain2Text Decoding Model Reveals the Neural Mechanisms of Visual Semantic Processing

ResearchDGX agent

arXiv:2503.22697v3 Announce Type: replace-cross Abstract: Decoding sensory experiences from neural activity to reconstruct human-perceived visual stimuli and semantic content remains a challenge in ne

CAPruner: Conceptual-Adjacent Scene Graph Pruner for Enhancing 3D Spatial Reasoning of Large Language Models

ResearchDGX agent

arXiv:2606.07529v1 Announce Type: cross Abstract: Large language models (LLMs) have recently been applied to 3D vision-language (3D-VL) tasks, which require spatial reasoning to identify target object

Collaborative Edge-to-Server Inference for Vision-Language Models

Local AiDGX agent

arXiv:2512.16349v2 Announce Type: replace-cross Abstract: We propose a collaborative edge-to-server inference framework for vision-language models (VLMs) that reduces communication cost while maintain

FAWAM: Force-Aware World Action Models for Closed-Loop Contact-Rich Manipulation

ApplicationsDGX agent

arXiv:2606.08555v1 Announce Type: new Abstract: Force signals provide critical interaction cues for contact-rich robotic manipulation. However, existing methods mostly use force as an additional obser

GNSS-FM: A Self-Supervised Foundation Model for Daily GNSS Displacement Time Series

ResearchDGX agent

arXiv:2606.07725v1 Announce Type: cross Abstract: Displacement time series from Global Navigation Satellite Systems (GNSS) are essential for a wide range of applications, including monitoring tectonic

Human-Centered Benchmarking of Driver Monitoring Models

SafetyDGX agent

arXiv:2606.08123v1 Announce Type: cross Abstract: Vision-based driver monitoring systems are increasingly deployed in safety-critical intelligent transportation settings, yet they are almost always co

Less Is More: Training-Free Acceleration Framework of 3D Diffusion Models for Low-Count PET Denoising via Global-Local Trajectory Reduction

Local AiDGX agent

arXiv:2606.08751v1 Announce Type: new Abstract: Accurate quantification and uptake measurement in PET are critical for assessing disease progression and supporting clinical decision-making. While high

Liberating LLM Capabilities in Full-Duplex Speech Models

ResearchDGX agent

arXiv:2606.07547v1 Announce Type: cross Abstract: Speech-based large language models are typically constrained to spoken replies, which limits their user-facing outputs to what can be verbalized and s

mllm-shap: A Shapley Value Explainability Platform for Text-Audio Multimodal Large Language Models

SafetyDGX agent

arXiv:2606.07531v1 Announce Type: cross Abstract: We introduce mllm-shap, an open-source Python framework designed to extend Shapley Value (SV) explainability from text-only Large Language Models to M

Modeling Components and Connections in Cyber-Physical Systems

ResearchDGX agent

arXiv:2606.09645v1 Announce Type: new Abstract: Text based configuration files for cyber-physical systems show the hierarchy of component modules well but often hide the details of connections and int

MotionVLA: Injecting Geometric Motion into Vision-Language-Action Model

ResearchDGX agent

arXiv:2606.08288v1 Announce Type: new Abstract: Vision-language-action (VLA) models increasingly condition robot policies on history, depth, or 4D features to resolve ambiguity in long-horizon manipul

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

SafetyDGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

SafetyDGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

SEF-CLGC at SemEval-2026 Task 11: Logical Notation Impact on Language Model Performance

SafetyDGX agent

arXiv:2606.09157v1 Announce Type: cross Abstract: This paper revisits our pipeline called Syllogistic Evaluation Framework-Common Logic Grammar Construction (SEF-CLGC). We combine formal logical notat

Sparrow: Sparse Rollout for Stable and Efficient Long-context RL of Large Language Models

SafetyDGX agent

arXiv:2606.08446v1 Announce Type: cross Abstract: Despite being powerful, reinforcement learning with verifiable rewards (RLVR) induces extremely long COT, making it computationally expensive. Since R

STELLAR: Spatio-Temporal Environmental Learning with Latent Alignment and Refinement for Long-Tailed Species Distribution Modeling

SafetyDGX agent

arXiv:2606.08484v1 Announce Type: cross Abstract: Joint Species Distribution Modeling (JSDM) is a key enabler for biodiversity monitoring and conservation planning. However, accurate JSDM faces two co

Super excited to announce that @arcee_ai is the first major American AI lab to replace AWS S3 with Hugging Face for ALL their models and dat…

IndustryDGX agent

Super excited to announce that @arcee_ai is the first major American AI lab to replace AWS S3 with Hugging Face for ALL their models and datasets, public AND private 🔥🔥🔥 Multi-million $ partnership to

Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structured Electronic Medical Records

ApplicationsDGX agent

arXiv:2606.08903v1 Announce Type: new Abstract: Synthetic healthcare data are widely proposed as privacy-preserving substitutes for real patient data, yet their evaluation remains dominated by statist

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

SafetyDGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

this is the biggest wake-up call to protect and nourish open source AI if you don't build out sovereign and independent models+infra closed …

IndustryDGX agent

this is the biggest wake-up call to protect and nourish open source AI if you don't build out sovereign and independent models+infra closed labs will patronize you to an insulting degree mythos will b

Zero Touch Predictive Orchestration: Automating Time-Series Models for the Cloud-Edge Continuum

ResearchDGX agent

arXiv:2606.09787v1 Announce Type: new Abstract: The Cloud-Edge Continuum (CEC) enables latency-critical applications by distributing resources to the far edge, but its extreme volatility makes proacti

8 Jun 2026

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.07107v1 Announce Type: new Abstract: Most vision-language-action (VLA) models map observations directly to actions without explicit intermediate planning, which limits performance on long-h

Coinbase : 'At Coinbase we're working hot on routing prompts to cheaper models where appropriate, & in some cases have been able to keep cos…

IndustryDGX agent

Coinbase : 'At Coinbase we're working hot on routing prompts to cheaper models where appropriate, & in some cases have been able to keep costs roughly flat, while token usage continues to grow exponen

DataEvolver: Automatic Data Preparation for Large Language Models through Multi-Level Self-Evolving

ResearchDGX agent

arXiv:2606.07001v1 Announce Type: cross Abstract: High-quality training data is essential to large language models (LLMs) and typically requires extensive and costly manual curation. Existing automati

FAIR-Calib: Frontier-Aware Instability-Reweighted Calibration for Post-Training Quantization of Diffusion Large Language Models

SafetyDGX agent

arXiv:2606.06547v1 Announce Type: cross Abstract: Diffusion Large Language Models (dLLMs) refine tokens iteratively but commit them irreversibly, leading to a 'stability lag' where early decisions rem

On orbital stabilization of a circular motion primitive for a dynamic extension of the Dubins car model

ResearchDGX agent

arXiv:2606.07449v1 Announce Type: cross Abstract: This paper addresses orbital stabilization of a circular motion primitive for a dynamic extension of the Dubins car model within a transverse-lineariz

One video, now made for every feed and format. Upload your existing video, choose your desired aspect ratio and watch our editing model, Ale…

IndustryDGX agent

One video, now made for every feed and format. Upload your existing video, choose your desired aspect ratio and watch our editing model, Aleph 2.0, fill in the rest of the scene as if you made it that

SS-TPT: Stability and Suitability-Guided Test-Time Prompt Tuning for Adversarially Robust Vision-Language Models

TutorialsDGX agent

arXiv:2606.06943v1 Announce Type: cross Abstract: Vision-language models (VLMs) such as CLIP achieve strong zero-shot recognition but remain highly fragile under adversarial perturbations. Recent test

Sycophantic Praise: Evaluating Excessive Praise in Language Models

SafetyDGX agent

arXiv:2606.07441v1 Announce Type: new Abstract: Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively

The Identity Trap in EEG Foundation Models: A Diagnostic Audit

ResearchDGX agent

arXiv:2606.06647v1 Announce Type: new Abstract: Objective. EEG foundation models (FMs) report strong accuracy on clinical resting-state EEG. However, high accuracy under subject-disjoint cross-validat

Train Models Faster with JAX and MaxText Using NVFP4 on NVIDIA Blackwell

HardwareDGX agent

This NVIDIA technical article addresses how low-bit mixed-precision pre-training accelerates large language model training, where optimizing numerical precision can reduce training time across distrib

You can find full model results and technical implementation details on our blog: https://cognition.ai/blog/frontier-code

AgentsDGX agent

Cognition AI published detailed model results and technical implementation information on their blog, likely covering performance benchmarks, architecture specifics, and methodology for their frontier

6 Jun 2026

a smarter alternative to 'always use plan mode': always frame your task as a question, so that the model is invited to push back and rate th…

ToolsDGX agent

a smarter alternative to 'always use plan mode': always frame your task as a question, so that the model is invited to push back and rate the quality of the idea/suggest alternatives, rather than blin

Adapting Diffusion Language Models for Lossless Pixel-Level Image Transmission

ResearchDGX agent

arXiv:2606.06273v1 Announce Type: cross Abstract: Lossless pixel-level image transmission is a fundamental regime beyond semantic communications, because exact recovery requires both accurate symbol p

Amortizing Federated Adaptation: Hypernetwork Driven LoRA for Personalized Foundation Models

SafetyDGX agent

arXiv:2606.06154v1 Announce Type: new Abstract: Federated fine-tuning of foundation models using Low-Rank Adaptation (LoRA) offers a communication efficient solution for distributed learning. However,

← Previous
1…210211212213214…1017
Next →