AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
Human
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
84,547 results
5 Aug 2026

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics

Local AiDGX agent

arXiv:2608.03291v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning improves large language model (LLM) performance while also providing an observable interface to the model's reasoning

The Transformer Revolution, Part 1: Dynamic Processing through Output- Weight Interconnections

ResearchDGX agent

arXiv:2608.03921v1 Announce Type: new Abstract: This paper offers a new interpretation of the Transformer during inference. Against the 'stochastic parrot' view that large language models merely repro

Thinking of buying more DRAM right now...

Model ReleasesDGX agent

So I'm looking at https://huggingface.co/unsloth/DeepSeek-V4-Flash-0731-GGUF and I realize my 128GB of DRAM just isn't cutting it for this (incredibly powerful) model. If only I had another 64GB, I th

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Third-party cyber evaluations involving OpenAI models

Model ReleasesDGX agent

Third-party cyber evaluations involving OpenAI models And another one. I had to create a accidental-cyberattacks tag to keep track of them all! This post from OpenAI covers both the UK AI Safety Insti

Tight Worst-Case Bounds for the Smallest Eigenvalue of ReLU NTK Gram Matrices

ResearchDGX agent

arXiv:2608.03368v1 Announce Type: new Abstract: For n unit vectors x_1,ldots,x_n in R^d, we study the continuous ReLU derivative Gram matrix H, whose entries are obtained by averaging pairwise gated i

TimeRLM: Recursive Language Models Enable Precise Anomaly Localization in Long-Context Time-Series

Model ReleasesDGX agent

arXiv:2608.03391v1 Announce Type: new Abstract: Precise anomaly localization over long-context time series is a crucial task in monitoring applications across clinical care, industrial operations, fin

Tired Actor: Fatigue-Informed Character Control

SafetyDGX agent

arXiv:2608.03528v1 Announce Type: new Abstract: Replicating human behavior with physics simulation has been a long-expected goal in character animation. Existing efforts have achieved impressive perfo

To Describe or Construct Statistical Learning Models Using the Category-theoretical Language

ApplicationsDGX agent

arXiv:2608.03706v1 Announce Type: new Abstract: Statistical learning is a fascinating field that has long been the mainstream of machine learning/artificial intelligence. A large number of results hav

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent…

AgentsDGX agent

To give some more context on what we are building with Daiwa Securities: During our technical verification phase, we integrated our AI agent technologies, specifically our AI Scientist and AB-MCTS fra

ToolLIFT: Lifting Tool-Specific Trajectories into Function-Level Graphs for Generalizable Tool Planning

AgentsDGX agent

arXiv:2608.03468v1 Announce Type: new Abstract: Historical tool-use trajectories provide valuable experience for large language model (LLM) agents to plan and coordinate tool usage. Existing approache

Topological Simplification in Predictive Coding Networks

ResearchDGX agent

arXiv:2608.02816v1 Announce Type: new Abstract: We study the topology of learned representations in predictive coding networks (PCNs), a neuro-inspired bidirectional architecture, using a quantitative

Toward Certified Functional Safety for Industrial Humanoid Robots: The Fail-Passive Gap and a Feasibility Study

SafetyDGX agent

arXiv:2608.02809v1 Announce Type: new Abstract: Industrial humanoid robots are constrained less by locomotion or manipulation capability than by the immaturity of functional safety certification for l

Toward Understanding the Transferability of Adversarial Suffixes in Large Language Models

ResearchDGX agent

arXiv:2510.22014v2 Announce Type: replace-cross Abstract: Discrete optimization-based jailbreaking attacks on large language models aim to generate short, nonsensical suffixes that, when appended onto

Toward Visual Grounding: A Survey

ResearchDGX agent

arXiv:2412.20206v4 Announce Type: replace Abstract: Visual Grounding, also known as Referring Expression Comprehension and Phrase Grounding, aims to ground the specific region(s) within the image(s) b

Towards a new paradigm of scientific discovery with socialized artificial intelligence

ResearchDGX agent

arXiv:2608.02775v1 Announce Type: new Abstract: Scientific discovery has advanced through successive transformations in the organization of knowledge. Observation and experimentation established the e

Towards Improving Sequential Decision-Making in LLM Agents via Experience Memory

AgentsDGX agent

arXiv:2608.03420v1 Announce Type: new Abstract: Large language models have improved substantially on single-shot reasoning tasks, but their performance in sequential decision-making is less well under

Towards Reliable and Reproducible Fetal Brain Biometry: A Deep Learning Approach Using MRI

ResearchDGX agent

arXiv:2608.03724v1 Announce Type: new Abstract: Fetal brain biometry is essential for quantitative assessment of brain development, supporting gestational age estimation, developmental monitoring, and

Towards Robust Tool Use in Agents via Experience-Driven Adaptive Guidance

AgentsDGX agent

arXiv:2608.03403v1 Announce Type: new Abstract: The performance bottleneck of agents is increasingly shifting from model capability to the robustness of their execution processes. Tools play a central

TQLite: Multi-LLM Jury Guided Distillation for Real-time MQM Translation Quality Evaluation

ApplicationsDGX agent

arXiv:2608.02975v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated impressive performance in MQM-based translation quality (TQ) evaluation, and recent advances in large r

Traceable Multi-Agent System for Knowledge-Based Forecasting

AgentsDGX agent

arXiv:2608.03339v1 Announce Type: new Abstract: Enterprise forecasting increasingly relies on autonomous agents that interpret documents, search for data, generate code, and revise models. While this

TraceCAD: Trace-Guided Repair for Agentic CAD Generation

AgentsDGX agent

arXiv:2608.03062v1 Announce Type: new Abstract: LLM-based CAD agents produce executable parametric programs, but their correction loops may lose evidence about satisfied requirements, faulty operation

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

Model ReleasesDGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

Track4Action: Distilling World-Centric 3D Tracker into Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.03727v1 Announce Type: new Abstract: Action labels tell a vision-language-action (VLA) policy which robot commands to imitate, but not how those commands change the 3D world. The aligned de

Training Documents Reranker with Search Rubrics for Deep Research Agent

AgentsDGX agent

arXiv:2608.03527v1 Announce Type: cross Abstract: Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically sele

Trajectory-Guided Forget-Recover Network for Continual LLM Unlearning

ApplicationsDGX agent

arXiv:2608.03123v1 Announce Type: cross Abstract: Machine unlearning aims to eliminate the influence of sensitive data on a model. In the real world, unlearning requests arrive continually, which give

Trajectory inference via Acceleration Matching

Model ReleasesDGX agent

arXiv:2608.03916v1 Announce Type: new Abstract: Trajectory inference is a fundamental problem in many scientific domains: given a collection of unpaired snapshots of observations at discrete time poin

Try npm i -g cline on Cline!👀

Model ReleasesDGX agent

Try npm i -g cline on Cline!👀 Qwen3.8-Max is now available in ClinePass, a subscription for ~5x discounted access. Use it on Cline CLI w/ $4.99 special promo: npm i -g cline This is currently the most

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

Model ReleasesDGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2608.04007v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) enables LLMs to solve complex tasks through iterative tool interactions. However, existing reinforcement learning meth

Two-Way Garment Transfer: Unified Diffusion Framework for Dressing and Undressing Synthesis

ResearchDGX agent

arXiv:2508.04551v2 Announce Type: replace Abstract: While recent advances in virtual try-on (VTON) have achieved realistic garment transfer to human subjects, its inverse task, virtual try-off (VTOFF)

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human…

TutorialsDGX agent

Two weeks ago, I resigned from OpenAI to join Conduit as a founding researcher, where we're training models to non-invasively read the human mind. I've written some thoughts about what telepathy could

UHP Detection: LVLMs have their Unique Hallucination Pattern in the Consistency Space

ResearchDGX agent

arXiv:2608.03817v1 Announce Type: cross Abstract: Large vision--language models (LVLMs) demonstrate strong multimodal reasoning capabilities but remain prone to hallucination, where model predictions

UL-UNAS: Ultra-Lightweight U-Nets for Real-Time Speech Enhancement via Network Architecture Search

ResearchDGX agent

arXiv:2503.00340v2 Announce Type: cross Abstract: Lightweight models are essential for real-time speech enhancement applications. In recent years, there has been a growing trend toward developing incr

Uncovering Spontaneous Physics Representations in In-Context Learning

ResearchDGX agent

arXiv:2508.12448v2 Announce Type: replace-cross Abstract: In-context learning (ICL) lets large language models (LLMs) solve new tasks from prompts alone, across an ever-widening range of domains, yet

Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection

Model ReleasesDGX agent

arXiv:2608.03627v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for automated fact-checking, yet their susceptibility to gender bias in this context remains underexp

Unexpected Crash-Looping of Multi-GPU Box: Caught by Hand-Scrutinized Logs - A Tale of Overridden Keep-Alive Policy and Eviction Thrashing

Local AiDGX agent

I've been freelancing for over a decade now, and I can't stress enough the importance of thorough investigation when dealing with strange software behaviors. Recently, I ran into an issue where my mul

UniEvo-RS: Omni-Prompt Unified Remote Sensing Segmentation with Representative Exemplar-Driven Prototype Evolution

ResearchDGX agent

arXiv:2608.03911v1 Announce Type: new Abstract: Prompt-driven vision-language models (VLMs) hold immense promise for accelerating dense remote sensing (RS) annotation, but static models suffer from se

Unified Visuomotor Targets: Supervising VLAs Beyond Physical Actions

SafetyDGX agent

arXiv:2608.03563v1 Announce Type: new Abstract: VLA models are trained to predict robot actions from visual and language observations. This is a natural choice, but it creates a mismatch: VLMs encode

UniGD: A Unified Generative-Discriminative Framework for Industrial Retrieval

ResearchDGX agent

arXiv:2608.03150v1 Announce Type: new Abstract: Generative retrieval (GR) is a promising paradigm for industrial search advertising, yet its deployment is constrained by strict relevance and latency r

UniNav: A Unified World-Action Diffusion Model for Visual Navigation

Model ReleasesDGX agent

arXiv:2608.03244v1 Announce Type: new Abstract: Image-goal visual navigation is a fundamental capability for embodied agents. Existing navigation policies efficiently predict waypoint trajectories but

UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations

ResearchDGX agent

arXiv:2604.14606v2 Announce Type: cross Abstract: Universal speech enhancement (USE) aims to restore speech signals from diverse distortions across multiple sampling rates. We propose UniPASE, an exte

UniWorld-Design: From Pixel Generation to Layer-Native Design

Model ReleasesDGX agent

arXiv:2608.03971v1 Announce Type: new Abstract: We introduce UniWorld-Design, a framework that redefines image generation from flat pixel synthesis to structured visual composition, with semantic RGBA

Unlocking the future of shared storage: Filestore on Colossus

Model ReleasesDGX agent

Today, enterprise storage must be as agile, elastic, and responsive as the workloads it supports. Filestore, Google Cloud’s first-party, secure, scalable NFS file service, can service a wide-range of

Unsupervised Adversarial Domain Adaptation for Uterine layer Segmentation: From Labeled Cine to Unlabeled Dynamic EPI MRI

ResearchDGX agent

arXiv:2608.03762v1 Announce Type: cross Abstract: Uterine peristalsis is a key physiological phenomenon responsible for various functions across the menstrual cycle, intimately linked to uterine wall

UNVaMP: Neural Knowledge Tracing with Variational Regularization of Latent Knowledge Dynamics

ApplicationsDGX agent

arXiv:2608.03811v1 Announce Type: new Abstract: We introduce the Unified Neural Variational Measurement of Proficiency (UNVaMP) architecture, a knowledge tracing method that integrates observed studen

UrbanAgent: A Tool-Augmented Agent for Cross-System Urban Tasks

Model ReleasesDGX agent

arXiv:2608.03018v1 Announce Type: new Abstract: Modern cities rely on an increasing number of digital services to operate, but residents' daily needs are still difficult to meet. Services are fragment

Utilize a nvidia gpu and amd gpu together for 2 different ai models?

Model ReleasesDGX agent

We run a local model instance in our company that the dev we hired built for us. We're a trade business and we want to further use our on hand hardware for it. The specs given we have is a 5090 gpu wi

V-FIND: Revealing the Intrinsic Forgery Knowledge Encoded in Video Forgery Detectors

ResearchDGX agent

arXiv:2608.03008v1 Announce Type: cross Abstract: As generated videos become increasingly realistic, reliable video forgery detection is increasingly important. Existing studies typically optimize and

ValueFormer: A Causal Transformer Value Function with Stage-Aware Labels for Semi-Autonomous Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.02958v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies trained by behavior cloning fail silently: from the action stream alone, a collapsing rollout looks much like on

Variational Approximated Restricted Maximum Likelihood Estimation for Spatial Data

ResearchDGX agent

arXiv:2604.07635v2 Announce Type: replace-cross Abstract: This research considers a scalable inference for spatial data modeled through Gaussian intrinsic conditional autoregressive (ICAR) structures.

Vectra AI launches Vectra AI Pro to feed AI agents better attack signals

Model ReleasesDGX agent

Threat detection company Vectra AI Inc. today launched Vectra AI Pro, a product built to give the artificial intelligence agents now working inside security operations centers a more reliable read on

Verifiable Memory: Learning Unified Memory Management with Local and Global Verifiers for Large Language Model Agents

Local AiDGX agent

arXiv:2608.03137v1 Announce Type: new Abstract: Large language model (LLM) agents must retain reusable information, control a bounded active context, and recover earlier evidence during long-horizon i

Verified Tool Calls Improve LLM Agent Reliability Under Non-Atomic Failures

AgentsDGX agent

arXiv:2608.02645v1 Announce Type: cross Abstract: Large Language Model (LLM) agents rely on external tools to perform multistage tasks. Existing agent frameworks typically assume that tool calls are a

Verifier-Guided Model Discovery for Physical Dynamical Systems with Pretrained Symbolic Transformers

Model ReleasesDGX agent

arXiv:2608.02662v1 Announce Type: cross Abstract: Reliable forecasting of nonlinear physical systems underpins scientific discovery and engineering decision-making. Yet high-fidelity simulations are p

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

Model ReleasesDGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

Very interesting to see @JeffDean's pitch deck. Just look at those open science and engineering problems. Lots to advance there with automat…

ResearchDGX agent

Very interesting to see @JeffDean's pitch deck. Just look at those open science and engineering problems. Lots to advance there with automated ML engineering. AI for science and engineering is just ge

VetScore: Risk-Weighted Fact Verification for Veterinary Long-Form QA with Citations

ResearchDGX agent

arXiv:2608.03675v1 Announce Type: new Abstract: Citation excerpts can be used to increase the reliability of generated outputs and their faithfulness to cited sources, which is especially important in

VIBE: A VAD-Informed Benchmark for Entity-Centered Affective Profiling of Large Language Model Outputs

Model ReleasesDGX agent

arXiv:2608.03810v1 Announce Type: cross Abstract: Large language models routinely describe socially salient targets, including political figures, countries, religions, organizations, historical events

VIBE: Vector Index Benchmark for Embeddings

Model ReleasesDGX agent

arXiv:2505.17810v2 Announce Type: replace Abstract: Approximate nearest neighbor (ANN) search is a performance-critical component of many machine learning pipelines, and rigorous benchmarking is essen

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

Model ReleasesDGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

← Previous
1…105106107108109…1410
Next →