AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,661 results
24 Jul 2026

Opus 5 now available in Hermes Agent

Model ReleasesDGX agent

Claude Opus 5 is now released in the Hermes Agent, a product of Nous Research and Teknium. Users can access the model through multiple gateways, including the Nous Portal, OpenRouter, and Anthropic Di

Our team just shipped Fugu-Ultra v1.1! 🐡 By dynamically orchestrating the latest frontier models, we pushed performance up by 7.9 points. W…

AgentsDGX agent

Our team just shipped Fugu-Ultra v1.1! 🐡 By dynamically orchestrating the latest frontier models, we pushed performance up by 7.9 points. We are now beating Fable 5 in complex coding and reasoning tas

Out of Sight, Still in Mind: Token Compression for Omni-LLMs

ResearchDGX agent

arXiv:2607.21179v1 Announce Type: new Abstract: The goal of this paper is to reduce the input token cost of Omni-modal large language models (Omni-LLMs) at inference time. Omni-LLMs reason jointly ove

Content type
AllBlogX PostPaperYouTubeRedditGitHub

Overcoming the Communication-Performance Tradeoff in LLM Pretraining

Local AiDGX agent

arXiv:2508.15706v3 Announce Type: replace Abstract: Communication-efficient distributed training algorithms (e.g., DiLoCo) have received considerable interest due to their benefits for training large

pAI-Econ-claude: A Gated Human-in-the-Loop Multi-Agent Architecture for AI-Assisted Economic Theory Development

Model ReleasesDGX agent

arXiv:2607.21268v1 Announce Type: cross Abstract: In many social-science research tasks, such as economics, LLM-based agents must produce outputs for which no cheap, task-complete, machine-readable co

[Paper] Statistically-Lossless Quantization of Large Language Models

Model ReleasesDGX agent

Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as GPTQ and AWQ achieve practical compression but

PATS: Policy-Aware Training Scaffolding for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.21419v1 Announce Type: new Abstract: In long-horizon LLM agent reinforcement learning, weak policies often repeat similar failures, producing uninformative rollout trajectories and limiting

PC-Edit: Prompt-Contrastive Region Discovery and Region-Guided Editing

ResearchDGX agent

arXiv:2607.21318v1 Announce Type: cross Abstract: Replacing an object with one that differs in category or shape requires complete source removal, natural target formation unconstrained by the source

People are using Minecraft farms as AI agent benchmarks

Model ReleasesDGX agent

Someone modelled sugarcane farming as an integer program. See, sugarcane only grows next to water. Water costs one tile and can feed at most four cane tiles. The layout therefore becomes a coverage pr

PersonaGesture: Single-Reference Co-Speech Gesture Personalization for Unseen Speakers

ResearchDGX agent

arXiv:2605.06064v2 Announce Type: replace Abstract: We propose PersonaGesture, a diffusion-based pipeline for single-reference co-speech gesture personalization of unseen speakers. Given target speech

PersonaTrail: Benchmarking Personalized Web Agents through Browsing Trails

Model ReleasesDGX agent

arXiv:2607.20482v1 Announce Type: new Abstract: Recent advances in large language models have enabled web agents to autonomously execute complex tasks. In practice, users frequently provide underspeci

Perspective Latents as an Architectural Condition for Causal Emergence in Active Inference Agents

SafetyDGX agent

arXiv:2607.20708v1 Announce Type: new Abstract: A recent line of work measures causal emergence in reinforcement learning agents through Integrated Information Decomposition, reporting that Phi_r grow

PhantomFill: When the Form Demands an Answer, Language Models Invent One

Model ReleasesDGX agent

arXiv:2607.20492v1 Announce Type: cross Abstract: Language models in production do not write prose. They fill forms: JSON fields, function arguments, extraction templates. We show that the form itself

Phonetic forced alignment for low-resource language varieties: Model training and evaluation on Chengdu Mandarin

SafetyDGX agent

arXiv:2607.21332v1 Announce Type: cross Abstract: Phonetic forced alignment is a key technique in phonetic research, yet existing alignment systems lack specialized models for low-resource language va

PhysCoRe: Physics-Corrected Residual World Models for Material-Aware Deformable Dynamics

ResearchDGX agent

arXiv:2607.20653v1 Announce Type: cross Abstract: Predicting how deformable objects evolve under robotic manipulation is a longstanding challenge. Existing approaches typically rely on per-object opti

Physics-Informed Deep Learning Model for Cross-Modality Super-Resolution in Fluorescence Microscopy

ResearchDGX agent

arXiv:2607.21190v1 Announce Type: new Abstract: Cross-modality image translation offers a route to super-resolution fluorescence microscopy from low-resolution images while reducing phototoxicity and

PILD: Physics-Informed Learning via Diffusion

SafetyDGX agent

arXiv:2601.21284v2 Announce Type: replace-cross Abstract: Diffusion models have emerged as powerful generative tools for modeling complex data distributions, yet their purely data-driven nature limits

Pipelined Gradient Coding

ResearchDGX agent

arXiv:2607.20739v1 Announce Type: cross Abstract: In large-scale machine learning, distributed training commonly involves multiple workers evaluating the gradients of the model on different dataset pa

PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses

Model ReleasesDGX agent

arXiv:2603.13026v2 Announce Type: replace Abstract: Prompt injection poses serious security risks to real-world LLM applications, particularly autonomous agents. Although many defenses have been propo

PlanE: Meta Planning of Data, Tuning, and Inference for Extractive-based LLMs

ResearchDGX agent

arXiv:2607.20470v1 Announce Type: new Abstract: Enhancing the task-specific capabilities of Large Language Models (LLMs) primarily requires substantial instruction-tuning datasets. However, the sheer

Position Bias is Hidden Behind Ceiling Effects: A Permutation Diagnostic for LLM Benchmarks

Model ReleasesDGX agent

arXiv:2607.20864v1 Announce Type: cross Abstract: Position bias in multiple-choice LLM evaluation is widely cited as a confound in capability comparisons, but published measurements rely on single ans

Position: Natural Language Should Not Fully Replace Formal Languages

ApplicationsDGX agent

arXiv:2607.20432v1 Announce Type: new Abstract: Recent advances in large language models and their widespread adoption have prompted claims that natural language could entirely replace formal language

Position: Stop Reactively Patching Your Model Every Time and Start Proactive Test-Driven AI Development

ResearchDGX agent

arXiv:2607.20532v1 Announce Type: cross Abstract: Many modern AI systems are designed to operate under diverse, open-ended, use-cases. To help generalize deployed systems, many deployed-system mainten

Post-Hoc Reasoning in Chain of Thought: Decoding and Steering Pre-Committed Answers

ResearchDGX agent

arXiv:2603.01437v2 Announce Type: replace Abstract: As chain of thought (CoT) has become central to scaling reasoning capabilities in large language models (LLMs), it has also emerged as a promising t

Preference Tuning as Spectral Update Reorganization

Model ReleasesDGX agent

arXiv:2607.20438v1 Announce Type: cross Abstract: Preference-based post-training is usually understood through endpoint behavior, yet the learned update that produces this behavior remains largely opa

PrefReward: Learning User Preference Matrix for Personalized Text Generation

TutorialsDGX agent

arXiv:2607.21067v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable ability in generating personalized content by leveraging user histories and contextual cues. H

Probabilistic Physics-Aware Machine Learning Predictions of Electric Truck Energy Consumption with Field Data

TutorialsDGX agent

arXiv:2607.19054v2 Announce Type: replace Abstract: In this work, we incorporate first principle physics into the construction of data-driven methods by considering a model that accounts for the diffe

Probabilistic Residual Learning for Online Recommendations

ResearchDGX agent

arXiv:2607.20863v1 Announce Type: cross Abstract: Modern recommender systems are typically based on deep learning (DL) models, where a dense encoder learns representations of users and items. As a res

Probably no LLM will ever achieve that, no matter how many data centers they build.

SafetyDGX agent

Probably no LLM will ever achieve that, no matter how many data centers they build. Lets compare Amazon Prime to AI: According to market research from Consumer Intelligence Research Partners (CIRP), t

ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning

Model ReleasesDGX agent

arXiv:2607.21022v1 Announce Type: new Abstract: Improving video captioning quality typically demands retraining large vision-language models, an expensive and often impractical requirement. Existing t

Profiling Lightweight Large Language Models

Model ReleasesDGX agent

arXiv:2607.20806v1 Announce Type: new Abstract: Lightweight large language models (LLMs) are increasingly being deployed locally on personal computers and are expected to play a growing role in resour

Progressive Cramming: Reliable Token Compression and What It Reveals

ResearchDGX agent

arXiv:2607.21231v1 Announce Type: new Abstract: Token cramming compresses sequences into learned embeddings with near-perfect reconstruction, but fixed token budgets and 99% accuracy thresholds leave

PromptPack: Scaling LLM Annotation Agents for Online Recommendation

Model ReleasesDGX agent

arXiv:2607.20528v1 Announce Type: new Abstract: Online recommendation platforms increasingly use Large Language Models (LLMs) to extract structured features from ad creatives. While deploying a single

QATMA: Quantization-Aware Training with Multimodal Alignment for Open-Vocabulary Object Detection

SafetyDGX agent

arXiv:2603.05964v3 Announce Type: replace Abstract: Quantizing open-vocabulary object detection (OVOD) models reduces their memory and computational costs, but extremely low-bit quantization severely

Quality-Aware Multimodal Fusion Reveals Implicit Identity in Valence-Arousal Features

TutorialsDGX agent

arXiv:2607.21347v1 Announce Type: new Abstract: Conventional face recognition relies on static appearance cues and degrades in unconstrained settings with expression variation, occlusion, and poor lig

QuantiBias: Benchmarking Quantization-Induced Bias in LLMs

Model ReleasesDGX agent

arXiv:2607.21063v1 Announce Type: new Abstract: Almost every large language model that reaches a broad audience is quantized: trained in full precision, then compressed for efficiency. This step is as

Quick Demo of the new Auto-Control feature in my Open-Source App that monitors stuff on your screen using local LLMs, so you don't have to :))

Local AiDGX agent

TLDR: This is a demo of my open-source app which now auto-controls itself so you can monitor your downloads, renders, progress bars, or whatever's on your screen and camera :) Hey r/ollama !! I'm deve

RadioTrace: Transmitter-Aware Diffusion for Radio Map Estimation without Deployment-Time Fine-Tuning

Local AiDGX agent

arXiv:2607.20909v1 Announce Type: cross Abstract: Radio map (RM) estimation aims to reconstruct the spatial distribution of wireless signal characteristics, such as received signal strength (RSS), fro

RE-AD: Real-Time Requirement Adherence for Data Labeling

Model ReleasesDGX agent

arXiv:2607.20455v1 Announce Type: cross Abstract: Human-annotated data remains fundamental to training frontier Large Language Models (LLMs). However, crowd-sourced annotations often suffer from quali

RealVDeblur: One-Step Diffusion for Generalizable Real-World Video Deblurring

ApplicationsDGX agent

arXiv:2607.20628v1 Announce Type: cross Abstract: Real-world video deblurring remains challenging due to diverse motion patterns, complex degradations, and the scarcity of realistic training data, yet

RECO: Region-Aware Compensation for Extrinsic Perturbations in Roadside 3D Detection

AgentsDGX agent

arXiv:2607.20947v1 Announce Type: new Abstract: In intelligent transportation systems, roadside 3D object detection provides wide-area perception crucial for traffic understanding, cooperative early w

Recurrent Sinusoidal INRs for Efficient High-Fidelity Representation

ResearchDGX agent

arXiv:2607.21485v1 Announce Type: new Abstract: We study sinusoidal recurrence as an iterative mechanism for harmonic spectral enrichment in implicit neural representations (INRs). Our analysis reveal

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers

SafetyDGX agent

arXiv:2607.21010v1 Announce Type: new Abstract: Zero-shot summarization using Large Language Models (LLMs) has significantly advanced the abstractive summarization task by producing coherent and fluen

REFACT: Adaptive Fact Restatement for Compact and Faithful Chain-of-Thought Reasoning

Local AiDGX agent

arXiv:2607.20833v1 Announce Type: new Abstract: Large language models increasingly rely on long-form reasoning for complex tasks, yet their reasoning traces may drift away from the supplied context wh

Refusal-Gated Decoding: Preserving Refusal Behavior Under High-Temperature Sampling

Model ReleasesDGX agent

arXiv:2607.20791v1 Announce Type: new Abstract: High-temperature sampling is one of the primary mechanisms for increasing diversity in LLMs. Recent advances in truncation-based sampling techniques hav

REGARD: Regional Affective Differences in Large Language Models

Model ReleasesDGX agent

arXiv:2607.20722v1 Announce Type: new Abstract: Large language models trained and aligned within different linguistic and regional ecosystems may frame the same political, cultural, and geopolitical e

Regularized Optimization on Grassmann Manifold: Theory, Algorithm and Applications

ApplicationsDGX agent

arXiv:2607.21039v1 Announce Type: new Abstract: Spectral methods are among the most widely used techniques for community detection, clustering, and graph learning. Their performance, however, critical

Regulating autonomous and agentic AI

SafetyDGX agent

arXiv:2607.21345v1 Announce Type: new Abstract: Regulating activities where regulatees use autonomous and agentic AI is challenging. Regulatory assumptions about regulatee knowledge and control no lon

“related incidents have been happening for a while” and OpenAI has no real solution in sight. Confirms my longstanding conjecture that curre…

SafetyDGX agent

“related incidents have been happening for a while” and OpenAI has no real solution in sight. Confirms my longstanding conjecture that current approaches cannot be made safe. An OpenAI staffer talked

Relative Value Learning

Model ReleasesDGX agent

arXiv:2607.21120v1 Announce Type: cross Abstract: In reinforcement learning, critics typically estimate absolute state values V(s), estimating how good a particular situation is in isolation. However,

Reliability-Aware LLM Alignment from Inconsistent Human Feedback

SafetyDGX agent

arXiv:2607.20515v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is critical for aligning Large Language Models (LLMs) with human preferences. However, its efficacy is

ReliableTableQA:How Much Supervision Does Reliability Annotation Need?

ApplicationsDGX agent

arXiv:2607.20537v1 Announce Type: cross Abstract: We introduce ReliableTableQA, a framework for training an LLM to annotate the statistical reliability of tabular QA results, not whether the query is

Replit shipped a lot this month. Talk to Agent with your voice, build from Claude or Slack, and never re-explain your stack to Agent again. …

Model ReleasesDGX agent

Replit released several major features this month, including a voice‑enabled Agent that lets users talk directly to the platform. Users can now build from Claude or Slack, and the updated Agent rememb

Representation Robustness Under Executable Reasoning Constraints in Large Language Models for Mathematical Problem Solving

Local AiDGX agent

arXiv:2607.20520v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evaluated on mathematical problem solving, yet prior work often treats representationally equivalent formu

Representative Sets in Propositional Abduction

ResearchDGX agent

arXiv:2607.21183v1 Announce Type: cross Abstract: The propositional abduction problem is a well-known form of non-monotonic reasoning where we are asked to find an explanation of a given manifestation

Representing Entity Importance in AI Knowledge Systems: A Dual-Signal Framework of Audience Evaluation and Structural Authority

SafetyDGX agent

arXiv:2607.20925v1 Announce Type: new Abstract: AI knowledge systems require representations of entity importance for retrieval, recommendation, evidence selection, and knowledge-intensive reasoning.

Respectfully, @davidsacks, I strongly disagree with your take, and I feel that you reached your conclusions without looking at the data, and…

SafetyDGX agent

Respectfully, @davidsacks, I strongly disagree with your take, and I feel that you reached your conclusions without looking at the data, and that your conclusions will give Americans false comfort. -

Response drift across frontier large language models

ResearchDGX agent

arXiv:2607.20454v1 Announce Type: cross Abstract: All frontier large language models (LLMs) exhibit response drift -- producing outputs that deviate from expert-validated references -- yet the magnitu

Rethinking Open-World Video Anomaly Detection: Diagnosing Definition Blindness

Local AiDGX agent

arXiv:2607.20780v1 Announce Type: new Abstract: Open-world video anomaly detection (OWVAD) is expected to detect events that match a user-specified definition of abnormality. This requirement is stron

Riemannian Deep Learning: Modules, Networks, and Geometries

ResearchDGX agent

arXiv:2607.19305v2 Announce Type: replace-cross Abstract: Deep neural networks on manifold-valued representations have attracted growing interest, but many basic components remain tied to specific man

← Previous
1…212213214215216…1412
Next →