AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
Human
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
85,201 results
16 Jul 2026

What Models Express, Suppress, and Resist: Auditing Open-Weight LLMs with Persona Vectors

AgentsDGX agent

arXiv:2607.13162v1 Announce Type: cross Abstract: What a language model will and will not do is largely set during post-training, but which behaviors it expresses, hides, or resists is not revealed by

What Your Model Threw Away and Why You'll Want It Back: Masking, Fingerprinting, and Privacy from Discarded Geometry

ResearchDGX agent

arXiv:2607.13046v1 Announce Type: new Abstract: We develop a framework for the information discarded by machine learning models whose inputs carry a Lie group action. Given a representation pi of a Li

When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for LLM Agents

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.11619v2 Announce Type: replace Abstract: Running the same LLM agent on identical inputs yields 2.3-4.2 distinct action sequences per 10 runs; this behavioral variance constitutes a training

When Audio Separation Hurts Zero-Shot ASR: Evaluating SAM-Audio with Whisper on Bengali and English Speech

ResearchDGX agent

arXiv:2603.04710v2 Announce Type: replace-cross Abstract: Recent advances in automatic speech recognition (ASR) and speech enhancement have strengthened the common belief that cleaner audio should lea

When Bots Join the Team: Bot Adoption and the Institutional Fabric of Open-Source Software Projects

AgentsDGX agent

arXiv:2607.13679v1 Announce Type: new Abstract: AI agents are joining human teams, raising a basic question: when an automated agent becomes a regular participant, does group organization strengthen o

When is the combined load identifiable from a stress-intensity profile? A coupled forward-inverse study on SIFBench finite-element data

ResearchDGX agent

arXiv:2607.13074v1 Announce Type: cross Abstract: This work studies the inverse problem of recovering the relative magnitudes of the tension, bending, and bearing loads acting on a crack from its stre

When T2I Synthetic Data Backfires: Amplified Privacy Risks in Real-Synthetic Mix Training

ResearchDGX agent

arXiv:2607.13541v1 Announce Type: cross Abstract: To overcome data scarcity and privacy constraints in data collection, it has become standard practice across academia and industry to augment real tra

When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs

Model ReleasesDGX agent

arXiv:2602.17659v2 Announce Type: replace Abstract: Vision-Language-Action models (VLAs) promise to ground language instructions in robot control, yet in practice often fail to faithfully follow langu

Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

SafetyDGX agent

arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pi

With Argus Eyes: Assessing Retrieval Gaps via Uncertainty Scoring to Detect and Remedy Retrieval Blind Spots

ResearchDGX agent

arXiv:2602.09616v2 Announce Type: replace-cross Abstract: Reliable retrieval-augmented generation (RAG) systems depend fundamentally on the retriever's ability to find relevant information. We show th

WNOJ-LIO: A White-Noise-on-Jerk Motion-Prior EKF for High-Dynamic LiDAR-IMU Fusion

AgentsDGX agent

arXiv:2607.13405v1 Announce Type: new Abstract: LiDAR-inertial odometry (LIO) is a key component of autonomous navigation, but high-dynamic driving exposes two coupled challenges: intra-scan motion di

Worlds in One Demo: A Synthetic Data Engine for Learning Open-World Mobile Manipulation

ApplicationsDGX agent

arXiv:2607.13154v1 Announce Type: new Abstract: Learning open-world mobile manipulation policies requires vast data to achieve spatial generalization, long-horizon robustness, and scene generalization

15 Jul 2026

1D-Bench: A Benchmark for Iterative UI Code Generation with Visual Feedback in Real-World

Model ReleasesDGX agent

arXiv:2602.18548v2 Announce Type: replace-cross Abstract: Design-to-code translates high-fidelity UI designs into executable front-end implementations, but progress remains hard to compare due to inco

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a …

Model ReleasesDGX agent

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a completed sale. You send a photo and a one-line clue, and Her

A Bearing-Strength Method for Motion Estimation of Unknown Energy Emitters

ApplicationsDGX agent

arXiv:2607.12515v1 Announce Type: new Abstract: This paper studies motion estimation of moving energy emitters using passive sensors. The emitters may be light, acoustic, or radio sources. While the b

A Behavioral State Vocabulary in Sony ERS-111 R-CODE

ResearchDGX agent

arXiv:2607.12115v1 Announce Type: new Abstract: This paper presents a corpus-level analysis of generated behavior diagrams derived from Sony's R-CODE sample distribution for the ERS-111 AIBO. Rather t

A Biomimetic Myoelectric Tentacle Prosthesis with Sensorless Object Detection and Vibrotactile Feedback

ResearchDGX agent

arXiv:2607.09807v2 Announce Type: replace Abstract: This paper presents the design and evaluation of a myoelectric tentacle-shaped prosthesis integrating electromyographic (EMG) control, sensorless ob

A Calibrated Multimodal Ensemble for Ambivalence/Hesitancy Recognition: System Description and Private-Test Submission Strategy

Model ReleasesDGX agent

arXiv:2607.12176v1 Announce Type: new Abstract: Ambivalence and hesitancy (A/H) undermine digital behaviour-change interventions, and recognizing them automatically from video is the goal of the ABAW

A closer look at improved intelligence in GPT-Live: the model can keep a conversation going while helping with multiple tasks at once, like …

Model ReleasesDGX agent

A closer look at improved intelligence in GPT-Live: the model can keep a conversation going while helping with multiple tasks at once, like checking flights, pulling up local weather, and shaping an i

A Comparative Analysis of Institutional and Course Generative AI Policies within Higher Education: Implications for Instruction in Computing Education

Model ReleasesDGX agent

arXiv:2607.12296v1 Announce Type: cross Abstract: With the increased use of generative AI (GenAI) applications such as ChatGPT, higher education institutions (HEIs) have released a range of guidelines

A hybrid analytical-PINN model for subsurface simulation of geothermal heat exchangers in heterogeneous underground

ResearchDGX agent

arXiv:2607.12271v1 Announce Type: new Abstract: In this paper, a parametric physics-informed neural network for solving the heterogeneous soil thermal problem with borehole heat exchangers (BHEs) as s

A JoLT for the KV Cache: Near-Lossless KV Cache Compression via Joint Tucker and JL-Residual Allocation for LLMs

Model ReleasesDGX agent

arXiv:2607.12550v1 Announce Type: cross Abstract: The key-value (KV) cache has become the dominant memory cost of transformer inference. It grows with batch size, context length, and depth, and at lon

A Learning-Rate-Gated Failure of GRPO in a Small Language and Vision-Language Model Web Agent: A Controlled Null and Its Mechanism

Local AiDGX agent

arXiv:2607.12640v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards, and Group Relative Policy Optimization (GRPO) in particular, is now run routinely on a supervised checkp

A Longitudinal Analysis of Public Discourse on AI Ethics in Education Using Twitter Data

ApplicationsDGX agent

arXiv:2607.12295v1 Announce Type: cross Abstract: The rapid integration of artificial intelligence (AI) and generative AI (GenAI) into education presents significant opportunities to enhance teaching

A model drop by Thinky 🚨🚨 Have been doing some early testing on the model for the past couple of days. Here are some of my findings 1. The…

AgentsDGX agent

A model drop by Thinky 🚨🚨 Have been doing some early testing on the model for the past couple of days. Here are some of my findings 1. The reasoning is sharp and concise! Always love to see models tha

A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study

Local AiDGX agent

arXiv:2607.12886v1 Announce Type: new Abstract: Clinical notes contain many of the signs and symptoms that bring patients to care, yet this information rarely reaches structured fields. Existing extra

A Neurosymbolic Approach to Natural Language Formalization and Verification

SafetyDGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

A Replicate-and-Quantize Strategy for Plug-and-Play Load Balancing of Sparse Mixture-of-Experts LLMs

ResearchDGX agent

arXiv:2602.19938v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts (SMoE) architectures are increasingly used to scale large language models efficiently, delivering strong accuracy under fi

A Shared Subcircuit Lets LLMs Count Down Across Tasks

Model ReleasesDGX agent

arXiv:2607.12279v1 Announce Type: new Abstract: Writing a sentence of exactly twelve words; ending a DNA sequence at the right codon; formatting an ASCII table. These are all tasks that language model

A Shortcut to Statistically Steady-State Turbulence with Flow Matching

ResearchDGX agent

arXiv:2607.13022v1 Announce Type: cross Abstract: Many nonlinear physical systems exhibit an initial transient phase in which perturbations grow before nonlinear interactions lead to a statistically s

A Threshold Exceedance Framework for CBRN Uplift Evaluation in Frontier Language Models

ResearchDGX agent

arXiv:2607.12200v1 Announce Type: new Abstract: As frontier language models advance, policymakers and model developers need methods for assessing whether model access materially increases a non-expert

AAAI-26 Dual Submissions: Novel Challenges

SafetyDGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

ABot-3DWorld 0: A Universal World Model to Explore Any 3D Space

ResearchDGX agent

arXiv:2607.11673v2 Announce Type: replace Abstract: We present ABot-3DWorld 0, a universal multimodal 3D world model that turns text, image, and video inputs into high-fidelity, explorable 3D worlds.

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory

Model ReleasesDGX agent

arXiv:2607.10350v1 Announce Type: cross Abstract: Recent VLM and VLA systems have improved robotic perception and action prediction, yet long-horizon embodied agents still require a general runtime la

ABot-N1: Toward a General Visual Language Navigation Foundation Model

Model ReleasesDGX agent

arXiv:2607.10383v2 Announce Type: replace-cross Abstract: Visual Language Navigation foundation models aim to unify deep reasoning for grounded spatial decisions with broad versatility for diverse emb

Accelerated Mixing Time of Randomized Hamiltonian Monte Carlo

ResearchDGX agent

arXiv:2607.12902v1 Announce Type: cross Abstract: We show the Randomized Hamiltonian Monte Carlo (RHMC) algorithm has accelerated mixing time guarantees for sampling from log-concave probability distr

Accelerating Masked Diffusion Large Language Models: A Survey of Efficient Inference Techniques

ResearchDGX agent

arXiv:2607.12829v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) offer a theoretical advantage in parallel generation over standard autoregressive models. However, parallel ge

Accepted Prefixes Are Not All You Need: A Negative Result on PEFT-Based Block-Diffusion Drafting

Model ReleasesDGX agent

arXiv:2607.12422v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive language model inference by using a cheap drafter to propose multiple future tokens and a target model t

Accuracy and Normalized Accuracy under Length Bias: Analysis, Guidelines, and a Bayesian Alternative

SafetyDGX agent

arXiv:2607.12767v1 Announce Type: new Abstract: Multiple-choice benchmarks that rank candidate completions by conditional log-probability suffer from a length bias: because log-probabilities sum over

ACID: Adaptive Caching for vIDeo generation

ResearchDGX agent

arXiv:2607.12358v1 Announce Type: new Abstract: Video diffusion models produce high-quality generations but remain slow at inference due to their sequential denoising procedure. Caching-based accelera

Action-Aware Generative Sequence Modeling for Short Video Recommendation

TutorialsDGX agent

arXiv:2604.25834v2 Announce Type: replace Abstract: With the rapid development of the Internet, users have increasingly higher expectations for the recommendation accuracy of online content consumptio

ACZ-GSeg: Adaptive Concentric Zone-based Two-stage Ground Segmentation for LiDAR Point Clouds

Local AiDGX agent

arXiv:2607.12110v1 Announce Type: new Abstract: Ground segmentation is a fundamental prerequisite for autonomous navigation, environmental perception, and object detection in ground mobile platforms.

AdaPCLA: Adaptive Prior-Calibrated Logit Adjustment for Long-Tailed Longitudinal EHR Generation

ApplicationsDGX agent

arXiv:2607.12645v1 Announce Type: new Abstract: Generative modeling of longitudinal Electronic Health Records is increasingly important for privacy-preserving research, yet standard autoregressive mod

Adaptive Compute in Latent World Models: When Depth Helps, Hurts, or Doesn't Matter

ResearchDGX agent

arXiv:2607.10203v2 Announce Type: replace-cross Abstract: Adaptive-compute world models -- early-exit or mixture-of-depths predictors that spend variable depth per step -- assume depth buys better pre

Adaptive Cross-Modal Fusion with Sparse Attention for Pedestrian Crossing Intention Prediction

Model ReleasesDGX agent

arXiv:2607.12293v1 Announce Type: new Abstract: Predicting pedestrian crossing intention is a safety-critical task for autonomous driving, yet existing approaches often rely on single-modal inputs or

Adaptive Testing for LLM Evaluation: A Psychometric Alternative to Static Benchmarks

Model ReleasesDGX agent

arXiv:2511.04689v3 Announce Type: replace-cross Abstract: Evaluating large language models (LLMs) typically requires thousands of benchmark items, making the process expensive, slow, and increasingly

Adversarial Attacks on Online Handwriting using Salience-based Temporal Editing

ResearchDGX agent

arXiv:2607.12500v1 Announce Type: cross Abstract: Deep learning models for online handwriting recognition have been shown effective and are increasingly deployed in practical applications. However, th

Affordance-Guided Diffusion Prior for 3D Hand Reconstruction

ResearchDGX agent

arXiv:2510.00506v2 Announce Type: replace Abstract: How can we reconstruct 3D hand poses when large portions of the hand are heavily occluded by itself or by objects? Humans often resolve such ambigui

Agent-Safety Evaluations as Load-Bearing Evidence: A Vendor-Neutral, Cross-Harness Reconstructability Metric

Model ReleasesDGX agent

arXiv:2607.12469v1 Announce Type: cross Abstract: Many agent-safety evaluation results are not yet load-bearing evidence: identical nominal outcomes (task success, attack success, monitor scores) may

AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

AgentsDGX agent

arXiv:2607.11098v2 Announce Type: replace-cross Abstract: Tool-using LLM agents are mostly evaluated assuming all tools work. When a tool times out, returns a week-stale value, or has its description

Agentic orchestration: Enterprise AI organizations have a deployment problem, not a platform problem — and most are calling chatbots agents

Model ReleasesDGX agent

Across 101 enterprises, agent orchestration is consolidating onto model-provider platforms — Anthropic’s Claude leads by a wide margin — chosen for the gravity of the underlying model and judged on re

Agentic Service-Oriented Computing: A Manifesto for the Next Frontier of Service-Oriented Computing

AgentsDGX agent

arXiv:2607.12619v1 Announce Type: new Abstract: The rapid emergence of LLM-powered autonomous and semi-autonomous agents is reshaping software systems from static, request-response components into goa

Agentic systems for breast cancer treatment recommendations

Model ReleasesDGX agent

arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning

Agentic vision: Building visual intelligence with Amazon Bedrock and MCP servers

AgentsDGX agent

In this post, we walk you through the Computer Vision MCP Server, which illustrates this approach, representing how AI systems can process visual information and make intelligent decisions through a s

Agents-A1-4B (Qwen3.7-4B ???) : Scaling the Horizon, Not the Parameters

Model ReleasesDGX agent

MODEL + GGUF : https://huggingface.co/InternScience/models?search=a1-4b Technical Report Benchmark Qwen3.5-4B Agents-A1-4B Qwen3.5 Qwen3.6 Nex-N2-mini Agents-A1 🧠 Dense Models (~4B) 🔀 MoE Models (35B-

Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Stateful Personal Agents

Model ReleasesDGX agent

arXiv:2607.10526v2 Announce Type: replace Abstract: Stateful personal agents increasingly maintain long-term user profiles, episodic memories, and reusable skills. This persistence turns conversationa

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to …

SafetyDGX agent

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to unlock a similar flywheel for safety, where today's models c

AI DevOps startup MyDecisive launches with $12M and open-source SmartHub

Model ReleasesDGX agent

Artificial intelligence DevOps startup MyDecisive formally launched today and announced 12 million in new funding to bring to market an open-source foundation for managing observability data and a com

Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results

AgentsDGX agent

arXiv:2607.11183v2 Announce Type: replace Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plau

Amsterdam-based Monumental, which develops autonomous robotics and software for the construction industry, raised a $32M Series B led by Khosla Ventures (Tamara Djurickovic/Tech.eu)

AgentsDGX agent

Tamara Djurickovic / Tech.eu: Amsterdam-based Monumental, which develops autonomous robotics and software for the construction industry, raised a $32M Series B led by Khosla Ventures — Monumental deve

← Previous
1…250251252253254…1421
Next →