AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,419
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,934
  • Model Releases23,909
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,419Total entries
1Added by human
88,418Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,638 results
1 Jun 2026

Simulation of collision avoidance behavior in crowd movement by data-driven approach

SafetyDGX agent

arXiv:2605.31210v1 Announce Type: cross Abstract: Crowd movement simulation is essential for pedestrian safety management and facility layout optimization. Data-driven models enhance trajectory predic

Synthetic Stimuli, Real Gains: Rethinking VLM Fine-Tuning Through Fully Controlled Data Generation

SafetyDGX agent

arXiv:2511.11440v3 Announce Type: replace-cross Abstract: Performance gains of Vision Language Models (VLMs) obtained by fine-tuning are generally based on ad hoc data collection and annotation of rea

TAGA: A Tangent-Based Reactive Approach for Socially Compliant Robot Navigation Around Human Groups

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2503.21168v3 Announce Type: replace Abstract: Robots navigating human-populated environments must avoid collisions while respecting the social structure of crowds, particularly the implicit boun

The Long-Term Effects of Data Selection in LLM Fine-Tuning

Local AiDGX agent

arXiv:2605.30537v1 Announce Type: new Abstract: Data selection is increasingly used to reduce the cost of large language model (LLM) fine-tuning, with recent methods prioritizing samples by current ut

The SuperActivator Mechanism: Transformers Concentrate Reliable Concept Signals in the Tail

ResearchDGX agent

arXiv:2512.05038v2 Announce Type: replace Abstract: Concept vectors aim to enhance model interpretability by linking internal representations with human-understandable semantics, but their practical u

TokTalk: Expressive Real-time Facial Animation from Audio-LLM Tokens

ResearchDGX agent

arXiv:2605.31294v1 Announce Type: new Abstract: Recent advances in Audio-LLMs like GPT-4o have ushered in an era of conversational interaction with language models. Conversational avatars however, sti

Understanding the Fundamental Design Decisions of Retrieval-Augmented Generation Systems

TutorialsDGX agent

arXiv:2411.19463v3 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a critical technique for enhancing large language model (LLM) capabilities. However, pract

Unlocking Fine-Grained Translation Quality Estimation in LRMs through Synergistically Evolving Implicit and Explicit Reasoning

ResearchDGX agent

arXiv:2605.31378v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) still struggle with fine-grained translation quality estimation (QE), even with long reasoning chains. We argue that LRMs

Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation

Model ReleasesDGX agent

arXiv:2605.30916v1 Announce Type: new Abstract: AI benchmarks have well-documented limitations, with prior work examining contamination, saturation, and construct underspecification. Aggregation has r

31 May 2026

How can I force Z-Image to create full-body portraits?

Local AiDGX agent

Z-Image is a fast, open-source image model from Alibaba Tongyi Lab that has gained popularity for its speed and image quality. Users on r/StableDiffusion likely discuss techniques for forcing the mode

// The Efficiency Frontier // Cool paper on context management. As agents reuse the same documents and histories across many turns, the chea…

Model ReleasesDGX agent

// The Efficiency Frontier // Cool paper on context management. As agents reuse the same documents and histories across many turns, the cheapest context strategy is not fixed. This work describes a pr

30 May 2026

Anthropic’s chance of being long-term profitable is greater than OpenAI’s. But still not huge.

Model ReleasesDGX agent

Gary Marcus argues that while Anthropic has better prospects for long-term profitability compared to OpenAI, the overall probability of either company achieving sustained profitability remains modest.

How we contain Claude across products

Model ReleasesDGX agent

How we contain Claude across products A complaint I often have about sandboxing products is that they are rarely thoroughly documented, and in the absence of detailed documentation it's hard to know h

The efficiency frontier! Where do you think GPT-5.6 will land?

Model ReleasesDGX agent

The efficiency frontier! Where do you think GPT-5.6 will land? Claude Opus 4.8 has landed on DeepSWE Bench, posting a 58% Pass@1 and taking #2 overall behind GPT-5.5. It continues a broader trend: sli

29 May 2026

A Novel Tensor Product-Based Neural Network for Solving Partial Differential Equations

Model ReleasesDGX agent

arXiv:2605.29688v1 Announce Type: new Abstract: This paper presents the Tensor Product Network (TPNet), a novel neural architecture for efficient and accurate function approximation and PDE solving. T

A Training-Time Diagnostic for Generalization via the Log-Alignment Ratio

Model ReleasesDGX agent

arXiv:2605.28975v1 Announce Type: new Abstract: We study the log-alignment ratio (LAR), a measure of parameter-activation alignment, introduced in parameterization theory. We reformulate it as the ove

A Triple-Modal Contrastive Learning Framework with Sequence, Graph, and 3D Features for Drug-Target Interaction Prediction

Model ReleasesDGX agent

arXiv:2605.29926v1 Announce Type: new Abstract: Accurate prediction of drug-target interactions (DTI) is critical for drug discovery. Existing methods often rely on single-modal representations (e.g.,

AlloyDB Hot Standby: Faster failovers, consistent performance

Model ReleasesDGX agent

AlloyDB for PostgreSQL is a fully managed, PostgreSQL-compatible database service designed for the most demanding enterprise workloads. It combines the best of PostgreSQL with the power of Google, del

Anti Mode-Collapse in Mean-Field Transformer via Auxiliary Variables

ResearchDGX agent

arXiv:2605.30229v1 Announce Type: new Abstract: We use a mean-field-based transformer model to theoretically investigate how auxiliary variables, such as positional encoding, prevent mode collapse of

Balancing Multimodal Learning through Label Space Reshaping

Model ReleasesDGX agent

arXiv:2605.28869v1 Announce Type: cross Abstract: Multimodal learning often suffers from modality imbalance, where modalities that converge faster dominate optimization while others remain undertraine

Bandit Algorithms for Deep Brain Stimulation

Model ReleasesDGX agent

arXiv:2601.12699v2 Announce Type: replace Abstract: Deep Brain Stimulation (DBS) is an effective treatment for Parkinson's disease, but conventional fixed-parameter stimulation can reduce battery life

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

Model ReleasesDGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

'Be My Cheese?': Cultural Nuance Benchmarking for Machine Translation in Multilingual LLMs

Model ReleasesDGX agent

arXiv:2602.04729v2 Announce Type: replace Abstract: We present a large-scale human evaluation benchmark for assessing cultural localisation in machine translation produced by state-of-the-art multilin

Building and Road Recognition in Dense Urban Informal Settlements: A Dataset and Benchmark

Model ReleasesDGX agent

arXiv:2605.29856v1 Announce Type: new Abstract: As a widespread form of informal settlements, urban villages present significant challenges for sustainable urban development and governance. Precise ma

BullingerDB: A Dataset for Handwritten Text Recognition and Writer Retrieval

Model ReleasesDGX agent

arXiv:2605.30235v1 Announce Type: new Abstract: We present BullingerDB, a large-scale benchmark dataset for historical document analysis based on the correspondence of Heinrich Bullinger (1504-1575).

CamC2V: Context-aware Controllable Video Generation

TutorialsDGX agent

arXiv:2504.06022v3 Announce Type: replace Abstract: Recently, image-to-video (I2V) diffusion models have demonstrated impressive scene understanding and generative quality, incorporating image conditi

Causal Intelligence for Constraint-Aware Intervention Design to Induce State Transitions

TutorialsDGX agent

arXiv:2605.29008v1 Announce Type: new Abstract: Driving a system from one state to another through targeted interventions is a fundamental challenge in science, yet most predictive models offer limite

Composing Non-Conjugate Factor Graphs with Closed-Form Variational Inference

Model ReleasesDGX agent

arXiv:2605.29467v1 Announce Type: cross Abstract: Stacking probabilistic building blocks into deeper architectures typically breaks closed-form inference. We show that closed-form inference can be pre

Convergence Theory for Iterative LLM-Based Neural Architecture Search: A Parametric Cross-Entropy Framework with Closed-Form Proxy Reliability

ResearchDGX agent

arXiv:2605.30103v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as generators in iterative neural architecture search (NAS), yet no formal convergence theory exists

CRB-Guided Framework Design and Resource Allocation for Indoor mmWave ISCC Systems

ResearchDGX agent

arXiv:2605.29939v1 Announce Type: cross Abstract: Integrated sensing, communication, and computation (ISCC) provides a promising framework for indoor human-centric applications. In these applications,

Deja View: Looping Transformers for Multi-View 3D Reconstruction

SafetyDGX agent

arXiv:2605.30215v1 Announce Type: new Abstract: Recent feed-forward 3D reconstruction transformers have scaled to over a billion parameters, following the broader trend of increasing model capacity in

DiScoFormer: Plug-In Density and Score Estimation with Transformers

TutorialsDGX agent

arXiv:2511.05924v3 Announce Type: replace Abstract: Estimating probability density and its score from samples remains a core problem in generative modeling, Bayesian inference, and kinetic theory. Exi

Dynamic Mixture of Progressive Parameter-Efficient Expert Library for Lifelong Robot Learning

Model ReleasesDGX agent

arXiv:2506.05985v3 Announce Type: replace Abstract: A generalist agent must continuously learn and adapt throughout its lifetime, achieving efficient forward transfer while minimizing catastrophic for

Echoes within the Reasoning: Stealthy and Effective Watermarking via Chain of Thought

ResearchDGX agent

arXiv:2605.28890v1 Announce Type: cross Abstract: Large Language Models with Chain-of-Thought reasoning capabilities represent valuable intellectual property, yet existing black-box watermarking metho

Eigen-Spike Emergence and Quadratic Equivalents for Conjugate Kernels on Nonlinearly Separable Data

ResearchDGX agent

arXiv:2605.29669v1 Announce Type: cross Abstract: Recent work in random matrix theory (RMT) has developed the notion of deterministic equivalents: typically linear surrogate models that approximate th

ElevenLabs launches Dubbing v2, which it says preserves the original speaker's emotion, tone, and pacing across 90+ languages while staying synced to content (ElevenLabs)

Model ReleasesDGX agent

ElevenLabs: ElevenLabs launches Dubbing v2, which it says preserves the original speaker's emotion, tone, and pacing across 90+ languages while staying synced to content — Today we're launching Dubbin

Ensemble Score Filtering for Real-Data Energy Consumption Forecast Correction

ResearchDGX agent

arXiv:2605.29072v1 Announce Type: new Abstract: Accurate estimation and forecasting of energy consumption are important for power-system operation, planning, and demand-side management. In practice, h

EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI

ResearchDGX agent

arXiv:2511.08949v2 Announce Type: replace Abstract: High-quality datasets are critical for training and evaluating reliable NLP models. In tasks like natural language inference (NLI), human label vari

F-RNG: Feed-Forward Relightable Neural Gaussians

ApplicationsDGX agent

arXiv:2605.25975v2 Announce Type: replace-cross Abstract: Capturing relightable 3D assets from real-world objects is a widely researched problem. Several per-scene optimization-based methods, based on

FakeVLM-R1: Internalizing Physical Laws via CoT for Synthetic Image Detection

SafetyDGX agent

arXiv:2605.30062v1 Announce Type: new Abstract: The development of generative artificial intelligence technologies has propelled the visual realism of synthetic images to an unprecedented level. Altho

FPLIER: Federated Pathway-Level Information Extractor

Model ReleasesDGX agent

arXiv:2605.29587v1 Announce Type: cross Abstract: In transcriptomics, gene-set-aware factorization methods such as the Pathway Level Information Extractor (PLIER) are most effective when trained on la

GenEraser: Generalizable Video Object Removal via Balanced Text-Mask Guidance and Decoupled Locator-Preserver

Model ReleasesDGX agent

arXiv:2605.30045v1 Announce Type: new Abstract: Video object removal frequently struggles to simultaneously eliminate target objects and their associated physical effects (e.g., smoke, reflections, li

GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling

AgentsDGX agent

arXiv:2605.28835v1 Announce Type: cross Abstract: Large Language Models (LLMs) extend their capabilities through function-calling (FC), which relies on training data with high quality, diversity, and

GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection

ResearchDGX agent

arXiv:2605.29539v1 Announce Type: cross Abstract: Vision-language foundation models have shown promising zero-shot generalization for Cross-Domain Few-Shot Object Detection (CD-FSOD). However, they fa

Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders

Model ReleasesDGX agent

arXiv:2605.30022v1 Announce Type: cross Abstract: Positional encoding (PE) underpins how permutation-invariant Transformers represent sequence order, yet how positional information is processed and st

Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation

ResearchDGX agent

arXiv:2602.08646v2 Announce Type: replace Abstract: We propose a gradient preconditioning method that makes reward-guided generation with one-step generative models both efficient and reliable. Test-t

GroundAct: Can LLM Agents Ground Actions in Environmental States?

Model ReleasesDGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing

Model ReleasesDGX agent

arXiv:2605.29532v1 Announce Type: cross Abstract: Exploratory GUI testing is a particularly demanding setting for MLLM agents: without predefined test scripts, an agent must autonomously navigate an a

Hallucination Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Semantic Caching

Model ReleasesDGX agent

arXiv:2605.29055v1 Announce Type: new Abstract: Hallucination remains a major reliability barrier for production LLM systems, particularly in multi-agent pipelines where unsupported claims can propaga

HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering

ResearchDGX agent

arXiv:2512.24562v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong question answering (QA) performance but can produce fluent answers unsupported by available evidence. Ex

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, …

Model ReleasesDGX agent

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer sit down on came

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that th…

Model ReleasesDGX agent

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that there would be a collapse of the AI bubble (which in my judgem

Improving Adversarial Robustness of Attribution via Implicit Regularization

Model ReleasesDGX agent

arXiv:2605.29983v1 Announce Type: cross Abstract: The adversarial robustness of attributions is a fundamental requirement for reliable explainability in deep learning, yet existing approaches typicall

In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration

ResearchDGX agent

arXiv:2510.00777v2 Announce Type: replace Abstract: LLM-generated drafts often contain subtle factual or logical errors, yet prior work shows that models struggle to reliably integrate multi-turn feed

InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents

Model ReleasesDGX agent

arXiv:2511.22884v2 Announce Type: replace Abstract: Data analysis has become an indispensable part of scientific research. To discover the latent knowledge and insights hidden within massive datasets,

KairosAgent: Agentic Time Series Forecasting with Fused Semantic Reasoning

AgentsDGX agent

arXiv:2605.30002v1 Announce Type: new Abstract: Cross-domain multimodal time series forecasting is a challenging task, requiring models to integrate precise numerical comprehension, cross-domain seman

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs

Model ReleasesDGX agent

arXiv:2605.29928v1 Announce Type: cross Abstract: As AI-generated and AI-assisted content floods online spaces, source labels attached to such content can distort human reasoning judgments, with downs

LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training

ResearchDGX agent

arXiv:2605.29888v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training has shown to improve reasoning in large language models (LLMs). However, there has been little exploration o

Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies

TutorialsDGX agent

arXiv:2605.29384v1 Announce Type: cross Abstract: We propose Latent Terms, a method revealing that models trained for dense retrieval, whether single- or multi-vector, learn representations that can t

Learning Representations from 3D Gaussian Splats

Model ReleasesDGX agent

arXiv:2605.29549v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) is a recent approach for scene rendering. Although primarily designed for view synthesis, its potential for scene understan

← Previous
1…544545546547548…1061
Next →