AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,483
  • Agents7,562
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,175
  • Local Ai4,939
  • Model Releases23,953
  • Research20,128
  • Safety13,373
  • Syntheses17
  • Tools1,677
  • Tutorials3,405

Source
HumanDGX agent

88,483Total entries
1Added by human
88,482Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,694 results
7 Aug 2026

Causal Episodic Memory for Feedback-Driven Agent Repair

Model ReleasesDGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

CodeGrep: An RL-Trained Retrieval Agent for LLM Coding Agents

Model ReleasesDGX agent

arXiv:2608.05886v1 Announce Type: cross Abstract: Modern LLM coding agents such as Claude Code and OpenHands share a common inefficiency: they spend much of their token budget finding the file to patc

Conditional Cognitive Biases in LLMs: How Biased User Turns Modulate In-Context Reasoning

Model ReleasesDGX agent

arXiv:2608.05166v1 Announce Type: new Abstract: We present an evaluation of cognitive bias expression in state-of-the-art instruction-tuned LLMs under realistic multi-turn interaction settings. Our wo

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Continuous-Time Piecewise-Linear Recurrent Neural Networks

TutorialsDGX agent

arXiv:2602.15649v2 Announce Type: replace Abstract: In dynamical systems reconstruction (DSR) we aim to recover the dynamical system (DS) underlying observed time series. Specifically, we aim to learn

Dense-Cast: A lightweight ensemble of deep learning architectures for precipitation nowcasting

ResearchDGX agent

arXiv:2608.06082v1 Announce Type: new Abstract: Proper short-term forecasting of precipitation is crucial in disaster management and preparedness. Nonetheless, the variability and nonlinearity of prec

DG-FedReuse: Proxy-Gradient-Gated Cached-Update Reuse with Matched Sparse Uplink Accounting

Local AiDGX agent

arXiv:2608.05358v1 Announce Type: new Abstract: Federated learning repeatedly incurs local optimization and model-update transmission. We study DG-FedReuse, a simulator-level mechanism that allows sel

EpiBench: Can LLMs Understand Epitopes for Antibody Drug Discovery?

Model ReleasesDGX agent

arXiv:2608.06022v1 Announce Type: new Abstract: Epitopes determine where antibodies bind antigens and shape downstream therapeutic properties such as functional blockade and escape resistance, making

Estimating time spent on work tasks

SafetyDGX agent

arXiv:2608.05172v1 Announce Type: cross Abstract: The task-based framework in economics models occupations as bundles of tasks. It is the standard lens for understanding how technology affects work: a

Evidence Lock Before Commitment: A Frozen Interface Degrades LLM-as-Judge Evaluation

Model ReleasesDGX agent

arXiv:2608.05353v1 Announce Type: new Abstract: LLM judges are often asked to extract criteria and evidence before choosing between candidate answers. This workflow assumes that the intermediate recor

EvReflection: Event-Driven Micro-Dynamics for Reflection Removal

Model ReleasesDGX agent

arXiv:2608.06184v1 Announce Type: new Abstract: Despite remarkable progress in reflection removal, current methods primarily exploit static image priors from a single frame and still suffer from sever

From Sports to Safety: Benchmarking Proactive Risk Inference in MLLMs

Model ReleasesDGX agent

arXiv:2608.05560v1 Announce Type: cross Abstract: Timely anticipation of physical hazards is essential for real-world safety, yet existing MLLM evaluations focus on harmful content or general risks, l

Got job as Director of AI and Systems development self-taught

Model ReleasesDGX agent

Hey everyone, I just wanted to share my journey here for some motivation. Three years ago, I saw the sudden spike in AI and realized it was the future of tech. My goal at the time was to be an indie g

Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

Model ReleasesDGX agent

arXiv:2608.06130v1 Announce Type: cross Abstract: AI agents performing cryptographic operations (signing Git commits, authenticating API calls, issuing certificates) currently store private keys in so

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

Model ReleasesDGX agent

arXiv:2608.06012v1 Announce Type: new Abstract: Search-agent rewards mix answer quality, citation grounding, tool cost, and anti-hacking terms; a high score therefore need not imply that cited evidenc

How Google Cloud detects, contains, and protects against emerging threats

Model ReleasesDGX agent

At Google Cloud, securing your data and business systems is our foundational commitment. We empower our customers with the tools, governance, and infrastructure needed to securely deploy workloads and

Hypothesis Testing with Conditional Queries: Learnability and the Value of Interaction

SafetyDGX agent

arXiv:2608.06262v1 Announce Type: new Abstract: Model evaluations may fix all tests before observing any responses or select later tests using earlier responses. We study this choice in a conditional-

IDperturb: Enhancing Variation in Synthetic Face Generation via Angular Perturbation

ApplicationsDGX agent

arXiv:2602.18831v2 Announce Type: replace Abstract: Synthetic data has emerged as a practical alternative to authentic face datasets for training face recognition (FR) systems, especially as privacy a

Inspecting Training Dynamics of Similarity Development in Supervised Vision Networks

SafetyDGX agent

arXiv:2505.21338v2 Announce Type: replace Abstract: For trustworthy and human-aware artificial intelligence, models should be evaluated beyond accuracy, among others through error predictability and s

Is Self-Pretraining really useful to improve diagnosis in medical Time Series?

ResearchDGX agent

arXiv:2608.06122v1 Announce Type: cross Abstract: Inspired by recent evidence that transformer architectures benefit from Self-PreTraining (SPT) on long-context benchmarks, we investigate whether simi

JoyAI-RA 0.5: Scaling Robot Manipulation Learning via Dual Action Alignment

Model ReleasesDGX agent

arXiv:2608.05674v1 Announce Type: new Abstract: Robot data is scarce, so generalist policies need to learn from heterogeneous sources, including human egocentric video, simulation, and real robots, wh

LC-GRPO: Bridging Train-Inference Gap for Flow-Based GRPO with Langevin Correction

SafetyDGX agent

arXiv:2608.05600v1 Announce Type: cross Abstract: Flow-based generative models are typically sampled by solving a deterministic ordinary differential equation (ODE), whereas online reinforcement learn

Mood Matters: How Syntactic Sensitivity Undermines Safety Alignment

SafetyDGX agent

arXiv:2608.05409v1 Announce Type: new Abstract: Large language models typically undergo post-training to align them with safety policies but there exist many sophisticated jailbreaks that sidestep est

Multi-Year Geospatial Reasoning using Interannually-Consistent Historical Predictions as a Free Input Modality

ResearchDGX agent

arXiv:2608.05979v1 Announce Type: new Abstract: Machine learning, and deep networks in particular, are increasingly used to derive higher-level Earth observation (EO) products such as annual land-cove

My issue with Artificial Analysis's 'intelligence index'

Model ReleasesDGX agent

I swear AA is not the bipartisan they so claim. An open source mode (Qwen 3.8 max) was number 1 on the agentic index, then they just so happen to launch 'v4.1.1' of their index in which they just adju

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

Model ReleasesDGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

PhaseCoder: Microphone Geometry-Agnostic Spatial Audio Understanding for Multimodal LLMs

Model ReleasesDGX agent

arXiv:2601.21124v2 Announce Type: replace-cross Abstract: Current multimodal LLMs process audio as a mono stream, ignoring the rich spatial information essential for embodied AI. Existing spatial audi

Positive-Unlabeled Preference Optimization For Chest X-ray Report Generation

ApplicationsDGX agent

arXiv:2608.05341v1 Announce Type: new Abstract: Vision-Language Models (VLMs) for radiology report generation are typically trained on retrospective clinical reports, which suffer from omission noise:

Project2Task: Graph-Guided Project-Level Planning for Autonomous Research

Model ReleasesDGX agent

arXiv:2608.05225v1 Announce Type: new Abstract: Research agents can increasingly search literature, propose hypotheses, generate code, run experiments, and draft manuscripts from a single topic. Howev

Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation

Model ReleasesDGX agent

arXiv:2608.06221v1 Announce Type: cross Abstract: Learning from demonstration (LfD) provides a developmental framework through which robots can develop motor skills by observing and imitating human dy

Safe Evolution with Circuit Anchors

SafetyDGX agent

arXiv:2608.05158v1 Announce Type: new Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential fun

SafeDivertor: Faithful Divertor Heat Flux Reconstruction from Macroscopic Plasma State Signals via Time-Frequency Prior Exploitation

Model ReleasesDGX agent

arXiv:2608.05669v1 Announce Type: cross Abstract: Divertor heat-flux analysis is essential for understanding plasma-wall interactions and protecting plasma-facing components in magnetic-confinement fu

StreamArena: Toward Continuous, Interactive, and Long-Horizon Agentic Streaming Video Understanding

Model ReleasesDGX agent

arXiv:2608.05703v1 Announce Type: new Abstract: Deploying autonomous multimodal agents in continuous, real-world environments requires them to ingest unbounded audio-visual streams and maintain hour-s

Text Generation: A Systematic Literature Review of Tasks, Evaluation, and Challenges

SafetyDGX agent

arXiv:2405.15604v4 Announce Type: replace Abstract: Text generation has become more accessible than ever, and the growing interest in these systems, especially those using large language models, has s

THBKG: A Temporal Biomedical Knowledge Graph for Decision-Aligned Clinical Advancement Prediction

Model ReleasesDGX agent

arXiv:2608.05982v1 Announce Type: new Abstract: Inadequate target--disease linkage accounts for 40--50% of Phase~II efficacy failures, so anticipating which programmes will advance would let sponsors

The cloud was always a cat. Qwen3.8-Max just saw it first. 😼☁️ Try it yourself!

Model ReleasesDGX agent

The cloud was always a cat. Qwen3.8-Max just saw it first. 😼☁️ Try it yourself! Qwen 3.8 Max is actually impressive Sent a sky pic to it along with Claude Opus 5, Kimi K3, and GPT 5.6 Sol, and asked t

Threshold-Based Early Stopping of Accumulations in Neural Networks with Binary Activation

Model ReleasesDGX agent

arXiv:2608.06177v1 Announce Type: new Abstract: Binary neural networks are very attractive for constrained deployment, enabling small footprint and low-power inference. For binary activations, the dot

Training-Free Token-Level Steering for LLM Personalized Co-Writing

ResearchDGX agent

arXiv:2608.06069v1 Announce Type: new Abstract: While Large Language Models (LLMs) show great promise for personalization, they often lack specialized domain knowledge. Conventional solutions like fin

Unified Agent: Managing Interactions across Devices

Model ReleasesDGX agent

arXiv:2608.05729v1 Announce Type: new Abstract: As capabilities rapidly increase, AI agents can move from running inside one app to acting across a user's devices over time. Yet existing agent systems

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing

Model ReleasesDGX agent

arXiv:2608.05485v1 Announce Type: new Abstract: Evaluating generated videos remains challenging because existing benchmarks rely on fixed evaluation content, cover only a subset of generation and edit

Visual Intention Grounding for Egocentric Assistants

Model ReleasesDGX agent

arXiv:2504.13621v2 Announce Type: replace Abstract: Visual grounding associates textual descriptions with objects in an image. Conventional methods target third-person image inputs and named object qu

When History Lies: Evaluating and Improving Tool Use under Misleading Multi-Turn Histories

Model ReleasesDGX agent

arXiv:2608.06057v1 Announce Type: new Abstract: Tool-calling agents infer task state from accumulated dialogue and tool traces. In persistent interactions, however, historical traces may remain struct

6 Aug 2026

A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction

ResearchDGX agent

arXiv:2608.04180v1 Announce Type: new Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, spar

A Modular Part-of-Speech Tagger for Scottish Gaelic using spaCy

ResearchDGX agent

arXiv:2608.04808v1 Announce Type: new Abstract: Part-of-speech tagging for low-resource languages remains challenging due to limited annotated data, especially for linguistically complex languages. Ga

AFD-Ledger: Deployment Provisioning for Attention--FFN Disaggregation

ResearchDGX agent

arXiv:2608.04502v1 Announce Type: cross Abstract: Attention--Feed-Forward Network (FFN) Disaggregation (AFD) is emerging as a promising architecture for serving Mixture-of-Experts (MoE) language model

ArborEnum: Decision Tree Rashomon Sets over Continuous Features

ResearchDGX agent

arXiv:2608.04310v1 Announce Type: new Abstract: The Rashomon effect describes the phenomenon that many models can achieve nearly equivalent performance on the same learning task, with significant rami

Attention-Only White-Box Transformer via LeJEPA-Based Self-Supervised Pretraining

Model ReleasesDGX agent

arXiv:2608.04213v1 Announce Type: new Abstract: Existing studies on self-supervised learning for white-box networks typically decouple the derivation of white-box networks via optimization algorithms

AudioScape-TTA: A Structured Soundscape Benchmark for Fine-Grained Text-to-Audio Evaluation

Model ReleasesDGX agent

arXiv:2608.04479v1 Announce Type: cross Abstract: Text-to-audio (TTA) generation has recently achieved remarkable progress in synthesizing realistic audio from natural language descriptions. However,

Beyond Semantic Equivalence: Logical Graphs for LLM Uncertainty Quantification

SafetyDGX agent

arXiv:2607.16868v2 Announce Type: replace Abstract: Large Language Models often produce confidently stated yet unreliable outputs, posing critical challenges for deployment in safety-sensitive applica

CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research

Model ReleasesDGX agent

arXiv:2605.12153v2 Announce Type: replace-cross Abstract: We present the Curated Industrial Developer Repository (CIDR), a large-scale dataset of real-world software repositories collected from indust

Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign diff…

AgentsDGX agent

Congrats to @mattrubens and the Roomote team on the launch. Builders can use Together AI as an inference provider in Roomote and assign different open models to coding, planning, vision, and review ac

Continual-Learning Physics-Informed Neural Networks for Parameterized Partial Differential Equations

Model ReleasesDGX agent

arXiv:2608.04778v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) incorporate governing equations into neural-network training and can approximate PDE solutions without requirin

Cooking beyond Frames: A Stereo Event Camera Dataset in the Kitchen

Model ReleasesDGX agent

arXiv:2608.04865v1 Announce Type: new Abstract: Event cameras, also known as neuromorphic cameras, have gained significant attention in recent years due to their high temporal resolution, high dynamic

Coupled Continuous-Discrete Generation for Scene Text Image Super-Resolution

Model ReleasesDGX agent

arXiv:2608.04525v1 Announce Type: new Abstract: Scene text image super-resolution (STISR) aims to recover visually plausible appearance while preserving character semantics from degraded inputs. Exist

DeepAmbigQA: Ambiguous Multi-hop Questions for Benchmarking LLM Answer Completeness

ResearchDGX agent

Large language models (LLMs) with integrated search tools show strong promise in open-domain question answering (QA), yet they often struggle to produce complete answer set to complex questions such a

Differential 6-DOF Pose Estimation with Provable First-Order Immunity to Camera Calibration Errors

Model ReleasesDGX agent

arXiv:2608.04673v1 Announce Type: new Abstract: Accurate six-degree-of-freedom (6-DOF) motion estimation is essential for robotic manipulation, autonomous systems, and structural displacement monitori

Dual 3090 setup: 400 pp t/s to 1600 pp t/s on Qwen 3.6 27B... with slightly lower tps.

Model ReleasesDGX agent

First of all, my setup: Ryzen 9 5950x DDR4 3200Mhz 64gb (2x32) Dual 3090s, no NVLINK Runtime: llama.cpp Nvidia Drivers 610 Windows 11 25H2 Qwen 3.6 27B Q8 I've been using llama-server with --split-mod

Dynamic Jailbreaking Attack

Model ReleasesDGX agent

arXiv:2510.02422v4 Announce Type: replace-cross Abstract: Existing gradient-based jailbreak attacks typically optimize a fixed-length adversarial suffix toward a predefined target response with a stat

Efficient Online Lexicographic Generalized Low-Rank Matrix Bandits

Model ReleasesDGX agent

arXiv:2608.04324v1 Announce Type: cross Abstract: This paper studies generalized low-rank matrix bandits with multiple prioritized objectives. At each round, the learner selects a matrix-valued arm an

EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement

Local AiDGX agent

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifie

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

SafetyDGX agent

arXiv:2608.04771v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) excel on complex tasks through long chain-of-thought (CoT) reasoning, but their lengthy intermediate steps cause severe ov

← Previous
1…507508509510511…1062
Next →