AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
26 Jun 2026

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current sta…

ApplicationsDGX agent

If you want to read an interesting AI thinking trace, try 'I want you to suggest two poems that you think apply very well to the current state of GenAI models like you. Don’t just pick popular poems a

Improved Bounds for Private and Robust Alignment

SafetyDGX agent

arXiv:2512.23816v2 Announce Type: replace-cross Abstract: In this paper, we study the private and robust alignment of language models from a theoretical perspective by establishing upper bounds on the

Kolmogorov Arnold networks (KAN) for aerodynamic prediction: a comparison with MLPs and GNNs

ResearchDGX agent

arXiv:2606.27126v1 Announce Type: new Abstract: Kolmogorov Arnold networks (KAN) have recently been introduced as a (deep) neural network architecture whose trainable parameters adapt the activation f

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

KRVF: A Source-Aware Semantic Voxel World Representation for Edge Mobile Manipulation

ResearchDGX agent

arXiv:2606.26321v1 Announce Type: new Abstract: Mobile manipulators need world models that are current, queryable, semantically meaningful, and usable under edge-compute constraints. This technical re

LCG: Long-Context Consistent Image Generation with Sparse Relational Attention

SafetyDGX agent

arXiv:2606.26171v1 Announce Type: cross Abstract: Recent image generation models achieve impressive quality in single-image synthesis, but often fail to maintain consistency across sequential outputs,

Learning to Explain Air Traffic Situation

AgentsDGX agent

arXiv:2502.10764v4 Announce Type: replace Abstract: Understanding how air traffic controllers construct a mental 'picture' of complex air traffic situations is crucial but remains a challenge due to t

LISA: Likelihood Score Alignment for Visual-condition Controllable Generation

SafetyDGX agent

arXiv:2606.27192v1 Announce Type: new Abstract: The prevalent dual-branch paradigm, i.e., training a side network to encode visual conditions and fusing its intermediate-layer features to a frozen pre

LogicIR: Logic Gate Networks for Image Restoration

ResearchDGX agent

arXiv:2606.26609v1 Announce Type: new Abstract: Image restoration aims to reconstruct high-quality images from degraded low-quality inputs. As the computational demands of image restoration models con

Multiscale Exit-Join Dynamics: Tactical Consensus and Strategic Coalition Formation

AgentsDGX agent

arXiv:2606.26139v1 Announce Type: cross Abstract: This paper develops a multiscale model of coalition formation in which strategic exit-and-join decisions are coupled with tactical consensus dynamics

oh and also...750 token/sec coming to 5.6 sol in july!

IndustryDGX agent

Sam Altman announced that OpenAI's inference speed will reach 750 tokens per second on GPUs with 5.6 TFLOPS in July. This represents a significant increase in throughput capability for OpenAI's models

PAMAE: Phase-Aware-MoE Action Experts Towards Reliable Flow-Matching Vision-Language-Action Policies

SafetyDGX agent

arXiv:2606.27144v1 Announce Type: new Abstract: Reliable action generation for multi-stage robotic manipulation remains challenging for Vision-Language-Action (VLA) models. While existing flow-matchin

Patronus AI grabs $50M in funding to stress-test AI agents in simulated environments

IndustryDGX agent

Fast-growing world model startup Patronus AI Inc. is priming itself for even more rapid growth after raising 50 million in Series B funding today. The round was led by Greenfield Partners and saw the

Semantic Early-Stopping for Iterative LLM Agent Loops

SafetyDGX agent

arXiv:2606.27009v1 Announce Type: new Abstract: Multi-agent large language model (LLM) loops, for example a Writer that drafts and a Critic that revises, are almost always terminated by a fixed iterat

something has definitely shifted in the past few weeks. seeing a huge uptick in large enterprises wanting to secure compute and post-train t…

ResearchDGX agent

something has definitely shifted in the past few weeks. seeing a huge uptick in large enterprises wanting to secure compute and post-train their own models in house, frequently on top of GLM-5.2. ever

Trace AI SDK 7 operations with LangSmith!

AgentsDGX agent

LangSmith enables tracing and monitoring of AI SDK 7 operations, allowing developers to observe and debug their AI applications built with the SDK. This tool provides visibility into model interaction

We’re getting together with @Zai_org for a FIFA World Cup watch party ⚽️ USA vs. the world on a big screen, food and drinks on us, and a fir…

AgentsDGX agent

We’re getting together with @Zai_org for a FIFA World Cup watch party ⚽️ USA vs. the world on a big screen, food and drinks on us, and a fireside chat on open-weight models and where agentic coding is

What to expect during RAISE Summit: Join theCUBE July 8-9

AgentsDGX agent

The artificial intelligence industry is entering a phase where the race for more capable models is taking a back seat to the economics of building and deploying AI infrastructure. This has resulted in

25 Jun 2026

A Bregman Perspective on Classification and Regression Trees

ResearchDGX agent

arXiv:2606.13984v2 Announce Type: replace-cross Abstract: Classification and Regression Trees (CART) constitute one of the most influential paradigms in statistical learning. Although a variety of imp

Agentic System as Compressor: Quantifying System Intelligence in Bits

AgentsDGX agent

arXiv:2606.25960v1 Announce Type: new Abstract: Large language models are turning from isolated predictors into agentic systems: they call tools, retrieve evidence, obey environment constraints, use v

An Empirical Study of Many-Shot In-Context Learning for Machine Translation of Low-Resource Languages

ResearchDGX agent

arXiv:2604.02596v3 Announce Type: replace Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks from a few examples, making it promising for languages underrepr

ASSCG: Just-Right Gating over Chattering for Fast-Slow LLM Planning in Autonomous Driving

AgentsDGX agent

arXiv:2606.25509v1 Announce Type: cross Abstract: Large language models (LLMs) can improve autonomous driving planning but are costly to query online, and existing fast-slow planners often rely on han

b9786

Local AiDGX agent

b9786 is a release of llama.cpp, a tool for LLM inference in C/C++ . Llama.cpp is a free and open-source tool that allows users to run AI models locally on Windows, Linux, and macOS . The b9786 releas

b9787

Local AiDGX agent

B9787 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference created by Georgi Gerganov. The llama.cpp project enables large language model inference with minimal setup

b9789

Local AiDGX agent

Release b9789 of llama.cpp includes a fix for quantizing mixture-of-experts models with MTP (multi-token prediction) . Binaries are provided for multiple platforms including macOS, Linux, Android, and

Black-Box Assisted Regression: Phase Transitions and Minimax Optimality

ResearchDGX agent

arXiv:2606.25743v1 Announce Type: new Abstract: Foundation models are often used as fixed black-box predictors for downstream tasks with limited labeled data, but their predictions may be biased and u

Breaking Data Symmetry is Needed For Generalization in Feature Learning Kernels

TutorialsDGX agent

arXiv:2604.00316v2 Announce Type: replace-cross Abstract: Grokking occurs when a model achieves high training accuracy but generalization to unseen test points happens long after that. This phenomenon

Build self-service AWS Health analytics to find actionable health insights with AI agents powered by Amazon Bedrock

AgentsDGX agent

In this post, we show you how to build Chaplin (Customer Health and Planned Lifecycle Intelligence Nexus), an open source solution that uses AI agents exposed through the Model Context Protocol (MCP)

But here's the punchline. Normalized to 90% cache hit rate: GLM-5.2 (Fireworks): 1.12/session Opus-4.7 (Anthropic): 2.14/session GLM is ~4…

ToolsDGX agent

This post compares the cost efficiency of Fireworks' GLM-5.2 model versus Anthropic's Opus across cached sessions, showing GLM-5.2 achieving approximately 4x lower cost per session at 90% cache hit ra

Chart of increasing revenue, pre the decline of token-maxxing, not looking at the immense costs/lack of profitability, not considering threa…

SafetyDGX agent

Chart of increasing revenue, pre the decline of token-maxxing, not looking at the immense costs/lack of profitability, not considering threats from open-source models, price-war dynamics etc. @noahopi

Choose your own

ResearchDGX agent

'Choose Your Own' likely refers to a framework or methodology from Nous Research for customizable AI decision-making or model selection processes. Without access to the specific post content, this ent

Chorus II: Cross-Request Sparsity Reuse for Efficient Image-to-Video Generation

ResearchDGX agent

arXiv:2606.25040v1 Announce Type: new Abstract: Serving diffusion models for image-to-video generation is computationally expensive, posing significant challenges for large-scale deployment. Real I2V

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

AgentsDGX agent

arXiv:2606.25332v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise for automated penetration testing, yet existing end-to-end black-box evaluations are highly susceptibl

Delving into Latent Spectral Biasing of Video VAEs for Superior Diffusability

Local AiDGX agent

arXiv:2512.05394v2 Announce Type: replace Abstract: Latent diffusion models pair VAEs with diffusion backbones, and the structure of VAE latents strongly influences the difficulty of diffusion trainin

Diagnosing and Mitigating Compounding Failures in Agentic Persuasion via Taxonomic Strategy Retrieval

AgentsDGX agent

arXiv:2606.24976v1 Announce Type: cross Abstract: Foundation-model agents in multi-step, open-ended environments frequently suffer from compounding errors, where early mistakes contaminate long-horizo

Domain-Specific Agents for Cherenkov Telescope Array Control Software and Gamma-Ray Data Analysis

AgentsDGX agent

arXiv:2510.01299v3 Announce Type: replace-cross Abstract: We present domain-adapted large language model agents designed to support Cherenkov Telescope Array operation and data analysis. The agents co

Dziri Voicebot: An End-to-End Low-Resource Speech-to-Speech Conversational System for Algerian Dialect

ResearchDGX agent

arXiv:2606.26003v1 Announce Type: new Abstract: Automatic speech and language technologies are still heavily biased toward high-resource languages, limiting their applicability to dialectal and low-re

Edges Before Embeddings: A Confidence-Aware Blur Gate for Vision-Language Pipelines

ApplicationsDGX agent

arXiv:2606.25838v1 Announce Type: new Abstract: Production vision pipelines silently degrade on blurry input, wasting compute on downstream OCR, retrieval, and vision-language model (VLM) calls that c

Efficient Analytic Uncertainty Quantification for Multi-Modal Regression

ResearchDGX agent

arXiv:2606.25188v1 Announce Type: new Abstract: Efficient uncertainty quantification (UQ) is essential for trustworthy large-scale learning. Existing UQ methods for regression tasks mainly operate und

ESMStereo: Enhanced ShuffleMixer Disparity Upsampling for Real-Time and Accurate Stereo Matching

Local AiDGX agent

arXiv:2506.21091v2 Announce Type: replace Abstract: Stereo matching has become an increasingly important component of modern autonomous systems. Developing deep learning-based stereo matching models t

FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation

SafetyDGX agent

arXiv:2606.26006v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are often constrained by the imitation ceiling imposed by sub-optimal data. While Reinforcement Learning (RL) fine-t

From Forecasting Leaderboards to Deployment Decisions: A Fail-Closed Certification Protocol

SafetyDGX agent

arXiv:2606.24996v1 Announce Type: new Abstract: Forecasting leaderboards rank models by predictive quality, but their winners are often read as deployment-ready top-1 advice. That reading can fail whe

GeoRanker: Distance-Aware Ranking for Worldwide Image Geolocalization

ResearchDGX agent

arXiv:2505.13731v4 Announce Type: replace Abstract: Worldwide image geolocalization-the task of predicting GPS coordinates from images taken anywhere on Earth-poses a fundamental challenge due to the

Graph-Based Phonetic Error Correction of Noisy ASR

Local AiDGX agent

arXiv:2606.24889v1 Announce Type: new Abstract: Automatic speech recognition (ASR) systems, despite low overall word error rates, produce residual lexical errors that disproportionately affect semanti

HEART: Coordination of Heterogeneous Expert Agents for Physically Grounded Robotic Task Planning

ApplicationsDGX agent

arXiv:2606.25404v1 Announce Type: new Abstract: Large Language Models (LLMs) can reason over complex instructions but often fail to satisfy the physical and spatial constraints required for robotic ta

Holographic Memory for Zero-Shot Compositional Reasoning in Knowledge Graphs: A Mechanistic Study of Where and Why It Fails

ResearchDGX agent

arXiv:2606.24948v1 Announce Type: new Abstract: Knowledge graph embedding (KGE) models predict single-hop links well but have no mechanism for zero-shot compositional queries: multi-hop questions whos

How KRAFTON Built PUBG Ally, a Co-Playable Character Powered by NVIDIA ACE

HardwareDGX agent

PUBG Ally is an interactive co-playable character for PUBG: BATTLEGROUNDS powered by an on-device small language model built with NVIDIA ACE technology, equipped to communicate and cooperate with play

Hybrid-IR: Dual-Path Hybrid Retrieval with Iterative Reasoning for Complex Medical Question Answering

ResearchDGX agent

arXiv:2606.25338v1 Announce Type: new Abstract: Large language models (LLMs) have shown promising performance across a wide range of biomedical applications, including medical question answering (QA),

I think this means you can collect ~10k hours just from open source datasets, which means basically anyone should be able to build a decent …

IndustryDGX agent

I think this means you can collect ~10k hours just from open source datasets, which means basically anyone should be able to build a decent robot foundation model: 500 from the new BitRobot dataset 50

Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors

ResearchDGX agent

arXiv:2606.25971v1 Announce Type: new Abstract: Modern neural network training relies on optimizers such as Adam and Muon which act on each weight matrix as a single object. Yet every weight matrix ca

Invisible to humans, visible to machines: a preregistered audit of Unicode fidelity across four biomedical bibliographic APIs

ResearchDGX agent

arXiv:2606.24897v1 Announce Type: cross Abstract: Biomedical text mining, scientometrics, and the construction of training corpora for biomedical large language models (LLMs) all assume that the abstr

Latent Space Analysis for Interpretable Uncertainty in Melanoma Classification

ResearchDGX agent

arXiv:2506.18414v3 Announce Type: replace Abstract: Melanoma is a highly aggressive skin cancer, making early and accurate diagnosis critical. While deep learning excels in skin lesion classification,

Learning Optimization Proxies for Sequential Contextual Stochastic Programs: An Order Fulfillment Application

SafetyDGX agent

arXiv:2606.25362v1 Announce Type: cross Abstract: Sequential contextual stochastic programs model real-time decision systems in which each time epoch commits to an action under uncertainty whose conse

LinStereo: Linear-Complexity Global Attention for Multi-Scale Iterative Stereo Matching

ApplicationsDGX agent

arXiv:2606.25437v1 Announce Type: new Abstract: Existing Vision Foundation Model (VFM)-based iterative stereo pipelines under-exploit three information pathways: multi-scale backbone features are coll

LLM-Based Scientific Peer Review: Methods, Benchmarks, and Reliability Challenges

SafetyDGX agent

arXiv:2606.25057v1 Announce Type: new Abstract: The rapid growth of scientific submissions has pushed traditional peer review toward its scalability limits, motivating the exploration of large languag

LLM Program Optimization via Retrieval Augmented Search

TutorialsDGX agent

arXiv:2501.18916v2 Announce Type: replace Abstract: Recent work has demonstrated the potential of large language models (LLMs) for program optimization, a key challenge in programming languages. We pr

Minimalist Preprocessing Approach for Image Synthesis Detection

ResearchDGX agent

arXiv:2606.25297v1 Announce Type: new Abstract: Generative models have significantly advanced image generation, resulting in synthesized images that are increasingly indistinguishable from authentic o

Mirendil raises $200M to speed up scientific research with AI

HardwareDGX agent

Mirendil Inc., a startup developing artificial intelligence models for scientists, has raised 200 million in funding at a 1 billion valuation. The seed round was led by Andreessen Horowitz. Mirendil s

MVTrack4Gen: Multi-View Point Tracking as Geometric Supervision for 4D Video Generation

ResearchDGX agent

arXiv:2606.26087v1 Announce Type: new Abstract: Synthesizing a novel-view video from a monocular reference video along a target camera trajectory requires both geometric consistency and motion fidelit

New MCP specification kills old risks but opens fresh attack surfaces, Akamai finds

AgentsDGX agent

A major overhaul of the Model Context Protocol due next month removes several longstanding protocol-level security risks but hands developers a fresh set of attack surfaces to defend, according to res

OPERA: Aligning Open-Ended Reasoning via Objective Perplexity-based Reinforcement Learning

SafetyDGX agent

arXiv:2606.25757v1 Announce Type: new Abstract: Reinforcement Learning (RL) has enabled LLMs to excel in objective reasoning tasks such as mathematics and code generation. However, applying RL to open

← Previous
1…744745746747748…1018
Next →