AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Model Releases

Flat Score, Amplified Failures: How the Error Budget Masks Damage in Quantized LLM Agents

DGX agent

arXiv:2607.27275v1 Announce Type: new Abstract: Post-training quantization to 4-bit weights is widely reported to be nearly lossless. We test this claim for multi-turn, tool-calling agents, where it n

model-releasesarxiv-cs-lg
31 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Generalization and Trade-off in Adversarial Training: An RKHS Perspective via Kernel Integral Operators

DGX agent

arXiv:2607.27995v1 Announce Type: cross Abstract: Adversarial training has emerged as a powerful approach for protecting models against adversarial attacks in a broad range of real-world applications.

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Gradient-free Task-Conditioned Retrieval for On-Device In-Context Learning

DGX agent

arXiv:2607.27766v1 Announce Type: new Abstract: On-device in-context learning (ICL) relies on pre-inference retrieval to select demonstrations for useful context before downstream model inference. Thi

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Hallucinations and Truth: A Comprehensive Accuracy Evaluation of RAG, LoRA and DoRA

DGX agent

arXiv:2502.10497v2 Announce Type: replace Abstract: Recent advancements in Generative AI have significantly improved the efficiency and adaptability of natural language processing (NLP) systems, parti

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

IFCMemoryBench: Evaluating Long-Term Memory of LLM-Based Agents in BIM Information Retrieval

DGX agent

arXiv:2607.26072v1 Announce Type: cross Abstract: Long-term memory is becoming a core capability of LLM-based agents, but existing evaluations largely test conversational recall in open-domain or pers

model-releasesarxiv-cs-ai
31 Jul 2026
Research

Metaphor Tracer: A Theory-Informed Analysis of Hidden States

DGX agent

arXiv:2607.28434v1 Announce Type: cross Abstract: What do a language model's hidden states say about the organization of a single text? From one forward pass, without training, we score every token po

researcharxiv-cs-cl
31 Jul 2026
Research

Meteosat Third Generation imagery improves CNN-based SSI retrieval

DGX agent

arXiv:2607.28093v1 Announce Type: cross Abstract: Accurate Surface Solar Irradiance (SSI) estimation is increasingly important for photovoltaic energy monitoring and forecasting. The recently introduc

researcharxiv-cs-lg
31 Jul 2026
Model Releases

Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?

DGX agent

Laguna 2.1 at NVFP4 Deepseek v4 at Q2 Inkling-Small at IQ3 Which models you guys running now ? How it compares to 122b? Upcoming in few days : Ling 3.0 124B (Could be new king) LongCat 69B A3B ( very

model-releasesr-localllama
31 Jul 2026
Safety

ODEWorld: A Continuous Predictive Architecture via Physical-Time Flow

DGX agent

arXiv:2607.27924v1 Announce Type: cross Abstract: In the physical world we inhabit, space and time are fundamentally continuous. However, existing machine learning paradigms for world modeling are lar

safetyarxiv-cs-cv
31 Jul 2026
Model Releases

On a joint simultaneous learning of relevant feature subsets and subspaces in regression-like problems

DGX agent

arXiv:2607.28080v1 Announce Type: cross Abstract: We extend a recently introduced Entropy-Optimal Manifold Clustering (EOMC) to allow for a joint simultaneous identification of subsets and subspaces o

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

OVEarth-Bench: Evaluating Category Breadth and Query Diversity for Open-Vocabulary Earth Observation

DGX agent

arXiv:2607.27278v1 Announce Type: new Abstract: Open-vocabulary Earth observation (EO) aims to localize geospatial concepts specified in natural language rather than a fixed label set. Existing benchm

model-releasesarxiv-cs-cv
31 Jul 2026
Research

Property-driven Causal Abstractions for Markov Decision Processes

DGX agent

arXiv:2607.26787v1 Announce Type: new Abstract: Markov Decision Processes (MDPs) are widely used as decision-making models, commonly specified over factored state spaces through state variables and th

researcharxiv-cs-ai
31 Jul 2026
Model Releases

Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents

DGX agent

arXiv:2607.28227v1 Announce Type: cross Abstract: GUI agents have the potential to become a general purpose executor over existing digital devices. To advance them toward real-world use, we envision a

model-releasesarxiv-cs-cv
31 Jul 2026
Agents

RoboBRIDGE: A Modular Framework for Bridging Policies to Robust Real-World Robotic Agents

DGX agent

arXiv:2607.27881v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have attracted growing interest as a scalable approach to robotic manipulation. While these models are effective act

agentsarxiv-cs-ro
31 Jul 2026
Model Releases

SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only Remediation

DGX agent

arXiv:2607.26313v1 Announce Type: cross Abstract: Agentic systems act, so a defect in the evidence they retrieve becomes a wrong action with a currency cost. The most dangerous enterprise defects are

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Scaling medical imaging report generation with multimodal reinforcement learning

DGX agent

arXiv:2601.17151v2 Announce Type: replace-cross Abstract: Frontier models have demonstrated remarkable capabilities in understanding and reasoning with natural-language text, but they still exhibit ma

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SemAnCorr: Semantic Anchored Correspondence for Zero-Shot Manipulation Skill Transfer

DGX agent

arXiv:2607.28382v1 Announce Type: new Abstract: Transferring manipulation skills across object instances that share functionality but differ in geometry remains a fundamental challenge in robot learni

model-releasesarxiv-cs-ro
31 Jul 2026
Model Releases

StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents

DGX agent

arXiv:2607.26314v1 Announce Type: cross Abstract: Stealth, the discipline of achieving an objective without revealing your presence, capabilities, or collected intelligence, is what separates sophisti

model-releasesarxiv-cs-ai
31 Jul 2026
Local Ai

When Should AI Follow? Task Structure and Joint Adaptation by Human and AI Agents

DGX agent

arXiv:2504.20903v4 Announce Type: replace-cross Abstract: How should organizations divide and sequence decision tasks between human and artificial agents? We develop a computational model of joint seq

local-aiarxiv-cs-ai
31 Jul 2026
Model Releases

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

DGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Advancing the price-performance frontier with GPT‑5.6

DGX agent

Advancing the price-performance frontier with GPT‑5.6 Huge price drop from OpenAI today: GPT-5.6 Terra got a 20% reduction, and GPT-5.6 Luna got a massive 80% drop. OpenAI credit 5.6 Sol with enabling

model-releasessimon-willison
30 Jul 2026
Model Releases

Aligning LLM-Simulated and Human Examinees for Psychometric Calibration: A Cognitive Diagnostic Profiling Approach

DGX agent

arXiv:2607.26317v1 Announce Type: cross Abstract: Psychometric calibration for educational tests typically requires costly human response data. Large language models (LLMs) simulated examinees offer a

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Batten Down Your Packages: Mitigation Guidance for Supply Chain Compromise

DGX agent

Written by: Kelli Vanderlee, Stuart Carrera For years, the cybersecurity industry's understanding of software supply chain compromise has been anchored by a few watershed events, including Russian cyb

model-releasesgoogle-cloud-ai
30 Jul 2026
Applications

ContactFlow: A video action conditioning that transfers across embodiments

DGX agent

arXiv:2607.26579v1 Announce Type: cross Abstract: World models offer a promising route toward robot planning by enabling agents to imagine and verify the consequences of actions before execution. Howe

applicationsarxiv-cs-cv
30 Jul 2026
Model Releases

Crossing-Free Probabilistic K-Line Forecasts Without Retraining

DGX agent

arXiv:2607.26792v1 Announce Type: cross Abstract: Probabilistic K-line forecasting describes uncertainty in four complementary prices, namely open--high--low--close (OHLC). However, it introduces two

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Diagnosing Fine-Grained Inconsistency Classification in Financial Disclosure Text

DGX agent

arXiv:2607.26368v1 Announce Type: new Abstract: Financial disclosures contain numerical claims, temporal statements, entity references, policy commitments, and risk descriptions that may conflict in q

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

EgoSafe: A First-Person Mobile-Captured Benchmark for Visual Safety Understanding

DGX agent

arXiv:2607.26518v1 Announce Type: new Abstract: Reliable visual safety understanding in real-world scenarios demands more than just object recognition; it requires causal reasoning under epistemic unc

model-releasesarxiv-cs-cv
30 Jul 2026
Model Releases

Enhancing Automated Machine Learning via Homogeneous Train-Test Splitting Methods

DGX agent

arXiv:2607.26625v1 Announce Type: new Abstract: Accurate model evaluation in machine learning depends critically on how datasets are split into training and testing subsets. Standard random splitting

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

FedWeave: Rethinking the Unit of Specialization in Heterogeneous Federated MoE-LoRA

DGX agent

arXiv:2607.26618v1 Announce Type: cross Abstract: Federated PEFT enables LLMs to collaboratively adapt to decentralized private data without sharing raw examples. However, task heterogeneity across cl

model-releasesarxiv-cs-cl
30 Jul 2026
Model Releases

Financial Volatility and Risk Forecasting Incorporating a Larger Number of Realized Measures

DGX agent

arXiv:2411.17136v2 Announce Type: replace-cross Abstract: Realised volatility has become increasingly prominent in volatility forecasting due to its ability to capture intraday price fluctuations. Wit

model-releasesarxiv-cs-lg
30 Jul 2026
Applications

Learning Dynamic User Personas from Implicit Interaction Streams via Iterative Refinement

DGX agent

arXiv:2607.26473v1 Announce Type: cross Abstract: Personalizing large language models (LLMs) to individual users is essential for improving user experience, yet existing approaches typically rely on e

applicationsarxiv-cs-cl
30 Jul 2026
Model Releases

LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving

DGX agent

arXiv:2607.26491v1 Announce Type: cross Abstract: The energy consumption of Large Language Model (LLM) serving is becoming a major system challenge as deployment scales, driven by hardware power and t

model-releasesarxiv-cs-lg
30 Jul 2026
Agents

Mitigating Compounding Error via Video Representation Regularization

DGX agent

arXiv:2607.27036v1 Announce Type: new Abstract: Video diffusion-based world models enable long autoregressive video generation for robotics, autonomous driving and simulation tasks, yet sliding-window

agentsarxiv-cs-cv
30 Jul 2026
Research

Prior Directions: Why GUI Grounding Gets Locked in the Past

DGX agent

arXiv:2607.26913v1 Announce Type: new Abstract: Vision-language models often use descriptions of earlier visual states to make decisions about the current scene. When the scene changes, stale language

researcharxiv-cs-cv
30 Jul 2026
Model Releases

Temporally Centered SIGReg Improves Multi-Task LeWorldModel Learning: From Analysis to Method

DGX agent

arXiv:2607.26924v1 Announce Type: new Abstract: Recent work on LeWorldModel (LeWM) has shown that the Sketched Isotropic Gaussian Regularizer (SIGReg) enables stable end-to-end world-model learning fr

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Transformers Can Learn Rules They've Never Seen: Proof of Computation Beyond Interpolation

DGX agent

arXiv:2603.17019v2 Announce Type: replace Abstract: A central question in the debate over large language models is whether transformers can learn rules they have never seen, or whether they can only i

model-releasesarxiv-cs-lg
30 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Agentic AI for Scientific Reasoning in Autonomous Quantum Sensing Experiments

DGX agent

arXiv:2607.25145v1 Announce Type: cross Abstract: We implement an agentic AI workflow built around a large language model (LLM) agent for autonomous experiments with nitrogen-vacancy (NV) centers in d

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

AI's Capability in Assisting Scientific Research in Physics, Astrophysics, and Cosmology II: Project Planning and Proposal Evaluation

DGX agent

arXiv:2607.25881v1 Announce Type: new Abstract: We investigate how well large language models (LLMs) can assist scientific project planning and proposal evaluation. One-page project plans were indepen

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

At-the-Roofline Sparse Tensor Contractions on Vector Processors for Transformer Inference

DGX agent

arXiv:2607.25504v1 Announce Type: cross Abstract: Fine-grained weight pruning and activation sparsification have emerged as effective approaches for reducing the compute and memory cost of inference f

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

A.X-K2 released

DGX agent

https://huggingface.co/skt/A.X-K2 https://huggingface.co/skt/A.X-K2-ALM https://huggingface.co/KRAFTON/A.X-K2-Raon-Speech-21B-A3B 688B-A33B + About South Korea's Soverign AI Foundation Model Project.

model-releasesr-localllama
29 Jul 2026
Model Releases

Bridging Compute- and Data-Optimal Pretraining

DGX agent

arXiv:2607.25271v1 Announce Type: cross Abstract: Classical compute-optimal scaling laws assume an unbounded supply of fresh pretraining data, yet pretraining is increasingly entering a regime in whic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CARE-MH: Towards Unified, Reproducible, and Comparable Evaluation of Mental Health LLMs

DGX agent

arXiv:2607.24754v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to provide mental health support, requiring reliable evaluation of safety, empathy, and therapeutic

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition

DGX agent

arXiv:2607.25294v1 Announce Type: cross Abstract: Real-world tasks often require models to learn from task-specific context rather than relying only on pre-trained knowledge. While recent work has hig

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization

DGX agent

arXiv:2607.25659v1 Announce Type: new Abstract: Rubric-based reinforcement learning enriches language model training by evaluating model outputs against explicit criteria. Yet in GRPO-style pipelines,

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution

DGX agent

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify no

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

DGX agent

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing

model-releasesarxiv-cs-ai
29 Jul 2026
← Previous
1…440441442443444…1371
Next →