AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,522 results
31 Jul 2026

Beyond KV Reconstruction: Functional Reconstruction for MLA Draft Models in Speculative Decoding

Model ReleasesDGX agent

arXiv:2607.27269v1 Announce Type: new Abstract: Multi-head latent attention (MLA) is increasingly important for long-context LLM inference because compact latent states replace the growing key-value (

BioPro: Towards Difference-Aware Gender Fairness for Vision-Language Models

SafetyDGX agent

arXiv:2512.00807v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) inherit significant social biases from their training data, notably in gender representation. Current fairness interve

Capturing Token Tendencies for Training-Free Token Pruning in Multimodal Large Language Models

Local AiDGX agent

arXiv:2607.28341v1 Announce Type: new Abstract: While visual token pruning is essential for efficient Multimodal Large Language Models (MLLMs), existing training-free methods suffer from a critical li

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CLAM: Continuous Latent Action Models for Robot Learning from Unlabeled Demonstrations

ApplicationsDGX agent

arXiv:2505.04999v2 Announce Type: replace-cross Abstract: Learning robot control policies from demonstrations typically requires action-labeled expert data, which is expensive to collect through teleo

Divergence Decoding: Training-Free Capability Fusion

Model ReleasesDGX agent

arXiv:2607.27248v1 Announce Type: cross Abstract: While large language models excel in reasoning, these generalists often lack knowledge for specialized scientific domains. Conversely, domain models~(

EgoGenesis: Egocentric World-Action Modeling with Online Anchored Projective Memory and Action-3D RoPE

SafetyDGX agent

arXiv:2607.28243v1 Announce Type: new Abstract: Egocentric video offers rich manipulation experience for embodied AI, yet collecting diverse egocentric data across scenes, objects, motions, and embodi

Interpreting learning dynamics of autoencoders: Transient scaling and emerging concepts of the Ising model

TutorialsDGX agent

arXiv:2607.10285v2 Announce Type: replace Abstract: We study how unsupervised autoencoders trained on microscopic spin configurations from the Ising model learn macroscopic, theory-relevant variables

Lightning OPD 2.0: Mitigating Style Bias in Cross-Teacher On-Policy Distillation for Large Reasoning Models

Model ReleasesDGX agent

arXiv:2607.28449v1 Announce Type: new Abstract: On-policy distillation (OPD) provides dense token-level supervision from a teacher, but its effectiveness can depend on teacher consistency, meaning tha

LightRot: A Light-Weighted Rotation Scheme and Architecture for Accurate Low-Bit Large Language Model Inference

Local AiDGX agent

arXiv:2607.27704v1 Announce Type: cross Abstract: As large language models (LLMs) continue to demonstrate exceptional capabilities across various domains, the challenge of achieving energy-efficient a

LM-GRASP: Instance-Specific Language Models for Combinatorial Construction via Online Imitation Learning

Model ReleasesDGX agent

arXiv:2607.28135v1 Announce Type: new Abstract: Machine learning for combinatorial optimization typically relies on neural constructors trained via reinforcement learning on large offline datasets for

OpenAI says its models now have more than 1B active users and are used by more than 2M businesses (Katherine Hamilton/Wall Street Journal)

IndustryDGX agent

Katherine Hamilton / Wall Street Journal: OpenAI says its models now have more than 1B active users and are used by more than 2M businesses — The announcement comes after OpenAI said earlier this week

Rethinking LLM-Judged Helpfulness as a Pedagogy Signal: A Pre-Registered Audit Across Tutor Models

Model ReleasesDGX agent

arXiv:2607.28128v1 Announce Type: new Abstract: LLM tutoring poses a measurement problem: can a general-purpose helpfulness rubric distinguish direct answer-giving from pedagogical guidance? We audit

30 Jul 2026

Challenges and proposed solutions in modeling multimodal medical data: A systematic review

TutorialsDGX agent

arXiv:2505.06945v5 Announce Type: replace Abstract: Multimodal data modeling has emerged as a powerful approach in clinical research, enabling the integration of diverse data types such as imaging, ge

DIRECT: Direct Decoding for Efficient and Aligned Sequence Labeling with Large Language Models

SafetyDGX agent

arXiv:2607.26891v1 Announce Type: new Abstract: Sequence labeling is a fine-grained information extraction task, yet existing large language model-based approaches suffer from insufficient domain alig

Explicit Kinematic Guidance from Analytic Concepts for Vision-Language-Action Models

TutorialsDGX agent

arXiv:2607.26513v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models rely mainly on 2D inputs, neglecting the rich object structural information and commonsense knowledge inhere

LumaGuide: Distribution Shaping for Training-Free HDR Generation in Diffusion Models

ResearchDGX agent

arXiv:2607.26237v1 Announce Type: new Abstract: Pretrained diffusion models generate realistic images but are constrained by the statistical biases of their training data, limiting their ability to pr

Object Detection for Autonomous Driving in Chinese Rural Scenes: An Experimental Study on Real-Synthetic Data Mixing and Model Evaluation

Local AiDGX agent

arXiv:2607.27058v1 Announce Type: new Abstract: Currently, autonomous driving object detection models face significant data scarcity and generalization challenges when navigating complex Chinese rural

Structurally Separated Uncertainty in Supervised Latent Variable Models

Model ReleasesDGX agent

arXiv:2602.11219v2 Announce Type: replace Abstract: Predictive uncertainty is commonly decomposed into epistemic and aleatoric components, but standard decompositions often produce strongly correlated

TPD: Temporal Prior Decoupling for Text-to-Video Diffusion Models

ResearchDGX agent

arXiv:2607.26706v1 Announce Type: new Abstract: Text-to-video diffusion models generate temporally coherent content from natural language, yet when a prompt describes an early scene that persists whil

29 Jul 2026

Balanced Soft mixture-of-expert model for Glaucoma Detection

TutorialsDGX agent

arXiv:2607.25324v1 Announce Type: cross Abstract: Glaucoma is a group of eye diseases that damage the optic nerve, often caused by elevated intraocular pressure. It is a leading cause of irreversible

Beyond Epistemia: Epistemic Schizologia and Large Language Models as Techno-Semiotic Machines

AgentsDGX agent

arXiv:2607.25620v1 Announce Type: new Abstract: Quattrociocchi and colleagues warn that the fluent outputs of large language models may allow linguistic plausibility to substitute for epistemic evalua

Desktop-Delta Bench: Do Computer-Use Models Understand Desktop GUI Transitions?

Model ReleasesDGX agent

arXiv:2607.26041v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly act through desktop GUIs to complete long-horizon tasks. Current benchmarks primarily measure end-task success o

Multi-Fidelity Learning with Shallow Recurrent Decoders for Multi-Physics Applications

Model ReleasesDGX agent

arXiv:2606.05202v2 Announce Type: replace-cross Abstract: In reactor physics, neutronics and multi-physics phenomena can be modelled at different fidelity levels. High-fidelity models based on the Bol

Online learning of subgrid-scale models for quasi-geostrophic turbulence in planetary interiors

ResearchDGX agent

arXiv:2511.14581v2 Announce Type: replace-cross Abstract: Machine learning approaches to subgrid-scale (SGS) modelling are now well established in atmospheric and oceanic applications. Among these, on

Universal Pansharpening Model

Model ReleasesDGX agent

arXiv:2603.03831v2 Announce Type: replace Abstract: Pansharpening generates the high-resolution multi-spectral (MS) image by integrating spatial details from a texture-rich panchromatic (PAN) image an

Wonder: Video World Model Done Better

ResearchDGX agent

arXiv:2607.26037v1 Announce Type: new Abstract: We present Wonder, a general-purpose video world model for real-time, camera-controllable world exploration. Given an image or a conditional video, Wond

28 Jul 2026

A Model for Imbalanced Label Aggregation: A Focus on Minority-Class Detection

ApplicationsDGX agent

arXiv:2607.24622v1 Announce Type: cross Abstract: We study imbalanced crowdsourcing with a focus on class-dependent annotator accuracy, a setting that, to the best of our knowledge, remains relatively

A New Kind of Adversarial Example: Measuring the Human-Model Gap, and Its Relationship to OOD Detection

ResearchDGX agent

arXiv:2607.22722v1 Announce Type: cross Abstract: Almost all adversarial attacks add an imperceptible perturbation to fool a model. We instead study the opposite: a large, clearly visible perturbation

Aligning Quantum Operators with Large Language Models

ResearchDGX agent

arXiv:2606.13811v2 Announce Type: replace-cross Abstract: Can Large Language Models (LLMs) understand and reason about quantum operators? Despite their remarkable capabilities in mathematics and symbo

An Efficient and Effective Evaluator for Text2SQL Models on Unseen and Unlabeled Data

ResearchDGX agent

arXiv:2603.07841v2 Announce Type: replace Abstract: Recent advances in large language models have strengthened Text2SQL systems that translate natural language questions into database queries. A persi

Beyond Block Boundaries: Multi-Block Editing for Diffusion Large Language Models

HardwareDGX agent

arXiv:2607.22663v1 Announce Type: new Abstract: Block diffusion has emerged as the dominant paradigm for scaling discrete diffusion language models (dLLMs), because decoding text in fixed-size blocks

CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph Databases

AgentsDGX agent

arXiv:2408.03910v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel in stand-alone code tasks like HumanEval and MBPP, but struggle with handling entire code repositories. Thi

Constraint-Bound Agnostic Bayesian Optimization: One Model for All Thresholds

Model ReleasesDGX agent

arXiv:2607.23448v1 Announce Type: cross Abstract: Expensive constrained optimization problems in real-world industry design often involve constraint thresholds that are difficult to determine in advan

Do Small Models Use the Law You Give Them? Context-Injected Fine-Tuning for Legal QA in Bangladesh

ApplicationsDGX agent

arXiv:2607.23446v1 Announce Type: cross Abstract: A small language model can receive the governing statutory provision and still answer incorrectly. We test whether fine-tuning on examples containing

Group Preference Collapse in Personalized Multimodal Large Language Models

ResearchDGX agent

arXiv:2607.22603v1 Announce Type: new Abstract: Personalized multimodal large language models (MLLMs) aim to generate user-specific responses, but existing methods mainly rely on profile-level informa

Hierarchical Grading in Large Language Models

ResearchDGX agent

arXiv:2607.22757v1 Announce Type: cross Abstract: We introduce Graded Large Language Models (GLLMs), an algebraic framework that equips the representation space of a transformer with a grading and pro

LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments

SafetyDGX agent

arXiv:2607.23969v1 Announce Type: new Abstract: World Action Models (WAMs) have emerged as a powerful paradigm for embodied intelligence, yet the prevailing reliance on pixel-level video generation cr

Let Me Look at You: Advanced Facial Expression Modeling for Conversational Speech Synthesis

ApplicationsDGX agent

arXiv:2607.24430v1 Announce Type: cross Abstract: Conversational Speech Synthesis is a fundamental component of human-computer interaction, aiming to generate contextually appropriate, expressive, and

Measuring Negative Campaigning across Languages with Large Language Models: A Study of 18 Million Tweets in 19 Countries

Model ReleasesDGX agent

arXiv:2507.17636v2 Announce Type: replace Abstract: Negative campaigning is a defining feature of electoral competition, yet comparative research on its drivers has remained limited by the high cost a

MedLoCoMo: A Long-Context Multi-Session Medical Dialogue Benchmark for Large Language Models

Model ReleasesDGX agent

arXiv:2607.22566v1 Announce Type: new Abstract: MedLoCoMo is a Medical Long-Context Memory benchmark for patient-specific clinical reasoning over multi-admission medical dialogue. Existing medical QA

ML-based Predictive Models for Power Consumption in Virtualised O-RANs

ResearchDGX agent

arXiv:2607.24256v1 Announce Type: cross Abstract: As communication networks adopt virtualized and disaggregated architectures, achieving energy efficiency has become increasingly important for both ec

N_0-TWAM: Scaling Tactile-Native World-Action Model for Contact-Rich Manipulation

ResearchDGX agent

arXiv:2607.23783v1 Announce Type: new Abstract: We present N_0-TWAM, a tactile-native world-action model for contact-rich manipulation that predicts both future vision and future contact. To our knowl

Omni-Prune: Query-Aware Unified Token Pruning for Efficient Omnimodal Large Language Models

HardwareDGX agent

arXiv:2607.23445v1 Announce Type: cross Abstract: Omnimodal large language models (OmniLLMs) are rapidly extending multimodal reasoning to cover synchronized audio and video. However, the resulting au

OmniCache: Multidimensional Hierarchical Feature Caching For Diffusion Models

ResearchDGX agent

arXiv:2607.23844v1 Announce Type: new Abstract: High-resolution image and video diffusion models, including SD3, FLUX, and recent video diffusion transformers, have substantially improved generative q

Reality Monitoring in Large Language Models: Self-Knowledge That Transforms with Conversation Memory

Model ReleasesDGX agent

arXiv:2607.23927v1 Announce Type: new Abstract: A conversational AI that cannot tell its own output from what a user said will treat its own mistakes as user-provided facts. In humans, this capacity i

Recently I've flipped from being bullish to being bearish about AI. I think I'm updating my bearishness to be more solidly bearish. Early th…

Model ReleasesDGX agent

Recently I've flipped from being bullish to being bearish about AI. I think I'm updating my bearishness to be more solidly bearish. Early thoughts (which I hope to be disproven in the next year or so,

SeT-Diff: Towards Semantic Foundation Models for HPC Telemetry and Time-Series

ApplicationsDGX agent

arXiv:2607.22548v1 Announce Type: new Abstract: Data centers and their compute nodes require accurate and flexible digital twins capable of modeling the complex interplay of workloads, environmental p

Similarity Is Not Logic: Factored Inference for Dual-Encoder Vision-Language Models

ResearchDGX agent

arXiv:2607.23052v1 Announce Type: cross Abstract: Dual-encoder vision-language models (VLMs) expose a similarity interface that enables zero-shot retrieval but fails compositional constraints: queries

Simulating Tenant Responses to Energy Policy Interventions with Transaction-Cost-Aware LLM Age

Model ReleasesDGX agent

arXiv:2607.24341v1 Announce Type: new Abstract: Recent studies use Large language models (LLMs) to simulate human opinions and decisions by prompting models with demographic, attitudinal, or persona-b

Systematic Analysis of Large Language Models and Transformer-Based Machine Translation for English-Tamil and Tamil-English Across Diverse Datasets

SafetyDGX agent

arXiv:2607.24515v1 Announce Type: new Abstract: The challenge of Machine Translation for low resource languages such as Tamil is primarily caused by the restricted amount of parallel data for these la

Temporal Context Reinstatement Drives Episodic-Like Order Memory in Long-Context Language Models

ResearchDGX agent

arXiv:2607.22575v1 Announce Type: new Abstract: Human episodic memory supports the retrieval of experiences that unfold over extended timescales, yet the computational mechanisms underlying this abili

UltraViT: Latency-Optimized On-device Vision Encoder for Large Vision-Language Models

Local AiDGX agent

arXiv:2607.23373v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) remain bottlenecked by massive computational footprints, precluding their deployment on resource-constrained edge d

UniGen-AR: Unifying Visual Generation with Auto-Regressive Modeling

ResearchDGX agent

arXiv:2607.24157v1 Announce Type: new Abstract: Modern computer vision pipelines remain fragmented, with tasks such as text-to-image generation, editing, restoration, and classical perception handled

27 Jul 2026

Developing and Validating the Spanish Version of the Large Language Models Dependency Scale (LLM-D12-SP)

ResearchDGX agent

arXiv:2607.22041v1 Announce Type: new Abstract: There is a growing need for reliable and culturally validated instruments to assess psychological dependency on large language models (LLMs), particular

Kimi K3 rivals Anthropic and OpenAI’s top models at a fraction of the cost. 3 / 1M input tokens. 15 / 1M output. $0.30 / 1M cached You can…

ToolsDGX agent

Kimi K3 rivals Anthropic and OpenAI’s top models at a fraction of the cost. 3 / 1M input tokens. 15 / 1M output. $0.30 / 1M cached You can now route your hardest reasoning to an open model without pay

Latent Interpolation Learning Using Diffusion Models for Cardiac Volume Reconstruction

ResearchDGX agent

arXiv:2508.13826v4 Announce Type: replace-cross Abstract: Cardiac Magnetic Resonance (CMR) imaging is a critical tool for diagnosing and managing cardiovascular disease, yet its utility is often limit

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can …

HardwareDGX agent

Make Kimi K3 yours with LoRA training on Fireworks Training a model this massive used to be a big project. With Fireworks Training, you can take the best open model in the world and efficiently tune i

Scaling Native Multimodal Pre-Training From Scratch

ResearchDGX agent

arXiv:2607.22043v1 Announce Type: new Abstract: Although large language models (LLMs) exhibit remarkable reasoning capabilities, their reliance on text-only pre-training restricts the perception of th

SurvDiff: A Diffusion Model for Generating Synthetic Data in Survival Analysis

ResearchDGX agent

arXiv:2509.22352v3 Announce Type: replace Abstract: Survival analysis is a cornerstone of clinical research by modeling time-to-event outcomes such as metastasis, disease relapse, or patient death. Un

Zero-Shot Mission-Level Evaluation for Aerial MLLM Agents

Model ReleasesDGX agent

arXiv:2607.22014v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are emerging as core reasoning modules for embodied agents, yet it remains unclear how well general-purpose m

← Previous
1…125126127128129…1009
Next →