AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,675 results
Model Releases

Qwen 27B is the 'DeepSeek moment' for open source. It matches the closed-source state-of-the-art from just a few months ago and runs on an R…

DGX agent

Qwen 27B is the 'DeepSeek moment' for open source. It matches the closed-source state-of-the-art from just a few months ago and runs on an RTX 5090. Without exaggeration, it’s a game changer. let that

model-releasesclem-delangue--x
18 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

R^3-Bench: LLMs Struggle with Resource-Rational Reasoning under Shared Budgets

DGX agent

arXiv:2608.16033v1 Announce Type: new Abstract: In cognitive science, resource rationality asks how an agent should allocate limited computation to maximize expected value. Most reasoning and agent be

safetyarxiv-cs-cl
18 Aug 2026
Model Releases

RamseyGadgets: A Graph Construction Dataset for LLMs

DGX agent

arXiv:2608.14999v1 Announce Type: cross Abstract: Constructing special graphs is an important task within graph theory and computer science. Many popular graph constructions are the result of a compre

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

RAPAC-DP: Response-Aligned Pending-Action Compensation for Diffusion Policies under Delayed Execution

DGX agent

arXiv:2608.15924v1 Announce Type: new Abstract: Cloud-side inference gives imitation-learning policies access to greater computational resources, but communication and computation delays can degrade c

model-releasesarxiv-cs-ro
18 Aug 2026
Model Releases

Readiness Barrier Functions: Forward-Invariant Control Authority for Overactuated Multirotor Allocation

DGX agent

arXiv:2608.16335v1 Announce Type: new Abstract: Allocation schemes that greedily maximize a readiness metric over the actuator fiber bundle of an overactuated multirotor produce commands that jump bet

model-releasesarxiv-cs-ro
18 Aug 2026
Safety

REFLEX: Reflexive Equilibrium Fixed-point Learning for Endogenous eXchanges

DGX agent

arXiv:2608.16155v1 Announce Type: new Abstract: In over-the-counter corporate bond markets, dealers compete for client trades by quoting bid and ask prices. Tighter quotes attract more business, but a

safetyarxiv-cs-lg
18 Aug 2026
Model Releases

Resource-Efficient QUBO Formulation for Anchored Currency Arbitrage

DGX agent

arXiv:2608.15889v1 Announce Type: cross Abstract: Currency arbitrage (CA) involves trading currencies in cycles to exploit discrepancies in market valuations. Quadratic unconstrained binary optimizati

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

Retrieval-guided Twin Fusion with Similarity-aware Contrast for Molecule-Text Alignment

DGX agent

arXiv:2608.16005v1 Announce Type: new Abstract: This paper studies the problem of molecule-text alignment, which aims to project molecules and their textual descriptions into a joint latent space for

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

Robot-Body-Aware Traversal Risk Graph Planning for Wheeled-Legged Robots in Complex Terrain

DGX agent

arXiv:2608.16433v1 Announce Type: new Abstract: Traversal Risk Graphs (TRGs) provide a compact, terrain-aware representation for global navigation, but native TRG costs are computed over circular node

model-releasesarxiv-cs-ro
18 Aug 2026
Safety

Robust structure from motion for aerial-ground images via detector-free feature matching and multi-view track refinement

DGX agent

arXiv:2608.15251v1 Announce Type: new Abstract: Integrated 3D reconstruction from aerial-ground images is essential for generating high-precision urban 3D models, yet severe variations in viewpoint, s

safetyarxiv-cs-cv
18 Aug 2026
Local Ai

RouteTS: Frequency-Time Routing for Time Series Forecasting

DGX agent

arXiv:2608.14682v1 Announce Type: new Abstract: Real-world time series inherently intertwine global periodic structures with localized non-stationary variations. Existing approaches process these hete

local-aiarxiv-cs-lg
18 Aug 2026
Safety

SCALE: State-Calibrated Latent Embeddings for JEPA Planning in the Right Geometry

DGX agent

arXiv:2608.16287v1 Announce Type: new Abstract: Joint-embedding predictive world models plan by scoring predicted terminal embeddings against a goal embedding using a cost defined on the representatio

safetyarxiv-cs-lg
18 Aug 2026
Model Releases

Second-Order Policy Effects as State Transitions: A Source-Linked Benchmark for Policy Simulation

DGX agent

arXiv:2608.15101v1 Announce Type: new Abstract: Policy evaluation often estimates direct benefits and costs while treating the institutional environment as fixed. In practice, a policy changes the sys

model-releasesarxiv-cs-ai
18 Aug 2026
Research

SEER: Long-Context Reasoning via Selective Visual-Text Compression

DGX agent

arXiv:2608.15962v1 Announce Type: new Abstract: Long-context reasoning remains computationally expensive for large language models due to the quadratic complexity of attention over text tokens. Visual

researcharxiv-cs-cl
18 Aug 2026
Safety

Semantic Bandits: In-Context Exploration-Exploitation is Biased by Semantic Priors

DGX agent

arXiv:2608.16707v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as decision-making agents in settings that require sophisticated environmental exploration. How

safetyarxiv-cs-ai
18 Aug 2026
Agents

Semantic Uncertainty-Guided Orchestration in Hierarchical Multi-Agent Systems

DGX agent

arXiv:2608.14707v1 Announce Type: new Abstract: As large language model (LLM)-based multi-agent systems become increasingly capable, coordinating agents under uncertainty becomes a fundamental challen

agentsarxiv-cs-ai
18 Aug 2026
Model Releases

Shape Operator PCA: Curvature-Aware Projections for Geometric Machine Learning

DGX agent

arXiv:2608.15313v1 Announce Type: cross Abstract: In this paper, we propose SHOPCA (Shape Operator-based Principal Component Analysis), a novel method for unsupervised metric learning and dimensionali

model-releasesarxiv-cs-ai
18 Aug 2026
Safety

SIGMA-Lane: Scale-pyramId Gated MAmba for Temporally Consistent Video Lane Detection

DGX agent

arXiv:2608.16338v1 Announce Type: cross Abstract: Video lane detection requires predictions that remain stable across frames, yet severe vehicle occlusions can break temporal cues. In streaming recurr

safetyarxiv-cs-ai
18 Aug 2026
Model Releases

Skill2Query: Exploiting Skill Structure to Generate Pseudo-Queries for Agent Skill Retrieval

DGX agent

arXiv:2608.16071v1 Announce Type: new Abstract: Pseudo-query generation can alleviate the supervision bottleneck for agent skill retrieval, but existing document-level approaches typically leave the r

model-releasesarxiv-cs-cl
18 Aug 2026
Local Ai

Solvable Sokoban Without a Solver via Diffusion

DGX agent

arXiv:2608.15958v1 Announce Type: new Abstract: Deciding whether a Sokoban puzzle is solvable is PSPACE-complete (Culberson, 1997): solutions can be exponentially long and there is no short certificat

local-aiarxiv-cs-ai
18 Aug 2026
Research

Spectral Saliency for Machine Unlearning

DGX agent

arXiv:2608.15548v1 Announce Type: cross Abstract: Machine unlearning (MU) aims to remove the influence of specific training data while preserving model utility. As the name suggests, MU can be viewed

researcharxiv-cs-ai
18 Aug 2026
Model Releases

STAGE: Controlled Objective Admission for Multi-Preference LLM Alignment

DGX agent

arXiv:2608.16553v1 Announce Type: new Abstract: Multi-preference alignment is often framed as scalarization: combine reward dimensions, then optimize. This leaves a temporal decision underspecified: w

model-releasesarxiv-cs-cl
18 Aug 2026
Safety

StructRL: Structured Action-Space Exploration for Flow-Based VLAs

DGX agent

arXiv:2608.15139v1 Announce Type: new Abstract: Flow-based Vision-Language-Action (VLA) models are now widely used for continuous robotic manipulation, and online reinforcement learning (RL) is emergi

safetyarxiv-cs-ro
18 Aug 2026
Tutorials

Supervising the Path to Fine Scales: GalerkinFlow for Scientific-Field and Image Super-Resolution

DGX agent

arXiv:2608.16546v1 Announce Type: new Abstract: Most super-resolution models learn from paired data by supervising only the final high-resolution output. This provides little control over how the pred

tutorialsarxiv-cs-cv
18 Aug 2026
Model Releases

Tac4Loco: Learning Spatiotemporal Plantar Pressure Representations for Humanoid Locomotion

DGX agent

arXiv:2608.15766v1 Announce Type: new Abstract: Humanoid robots are expected to traverse complex terrains, where the plantar support may vary dramatically due to foot placement errors, ground properti

model-releasesarxiv-cs-ro
18 Aug 2026
Model Releases

The Benchmark Trap: Structures of Power and Injustice in AI Evaluations

DGX agent

arXiv:2608.15326v1 Announce Type: new Abstract: Artificial intelligence (AI) benchmarks are not neutral tools of evaluation but socio-technical artefacts that shape competition, power, and research pr

model-releasesarxiv-cs-ai
18 Aug 2026
Research

The Limits of Binding in Dual Encoders

DGX agent

arXiv:2608.15971v1 Announce Type: cross Abstract: Dual-encoder models such as CLIP score an image-caption pair by a single inner product of two independently computed unit vectors, and fail at binding

researcharxiv-cs-cl
18 Aug 2026
Model Releases

The Recall Trap: A Recall-Maximizing Retriever Configuration Reduces Issue Resolution in Fixed-Budget Code Context

DGX agent

arXiv:2608.14838v1 Announce Type: cross Abstract: Retrieval components for code assistants are tuned against retrieval metrics: a configuration that raises recall@k is adopted, and downstream task suc

model-releasesarxiv-cs-cl
18 Aug 2026
Research

The Right Prior for the Right Deformation: Rethinking Continuous Deformable Image Registration

DGX agent

arXiv:2608.16146v1 Announce Type: new Abstract: Deformable image registration models implicitly encode deformation priors through their parametrization and optimization. In this work, we conduct a val

researcharxiv-cs-cv
18 Aug 2026
Research

Thinking Outside the (Gray) Box: A Context-Based Score for Assessing Value and Originality in Neural Text Generation

DGX agent

arXiv:2502.13207v4 Announce Type: replace-cross Abstract: Despite the increasing use of large language models for creative tasks, their outputs often lack diversity. Common solutions, such as sampling

researcharxiv-cs-ai
18 Aug 2026
Tutorials

Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs

DGX agent

arXiv:2603.06697v2 Announce Type: replace-cross Abstract: Vision--language models (VLMs) process images as visual tokens, yet their intermediate reasoning is often carried out in text, which can be su

tutorialsarxiv-cs-ai
18 Aug 2026
Research

TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability

DGX agent

arXiv:2405.17678v2 Announce Type: replace-cross Abstract: Achieving zero-shot adversarial robustness without sacrificing generalization remains challenging for foundation models such as CLIP, especial

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Token Distribution versus Data Volume: Domain Balancing in Multi-Domain Meeting Summarisation

DGX agent

arXiv:2608.15935v1 Announce Type: new Abstract: Jointly fine-tuning an LLM on meeting-summarisation corpora of widely varying size raises a question that prior work leaves confounded: when a domain-ba

model-releasesarxiv-cs-cl
18 Aug 2026
Safety

Towards a theory of inference-time alignment with unknown rewards

DGX agent

arXiv:2608.15402v1 Announce Type: new Abstract: Generative model alignment has received broad interest, and significant progress has been made in supervised fine-tuning and inference-time computation.

safetyarxiv-cs-lg
18 Aug 2026
Tutorials

Towards Reasonable Molecular Structure Elucidation from Infrared Spectroscopy with Chemical Feedback

DGX agent

arXiv:2608.16082v1 Announce Type: new Abstract: Infrared (IR) spectra provide characteristic signals of molecular structure, which are often interpreted by experts via functional-group identification

tutorialsarxiv-cs-lg
18 Aug 2026
Model Releases

Towards Zero-Shot Domain Generalization for ID Cards Presentation Attack Detection

DGX agent

arXiv:2608.16591v1 Announce Type: new Abstract: Presentation-Attack Detection (PAD) for national ID cards is limited by the lack of publicly available genuine samples, making it difficult for systems

model-releasesarxiv-cs-cv
18 Aug 2026
Model Releases

Training and Evaluating Ethical Reinforcement Learning Agents on Per-Episode Distributions

DGX agent

arXiv:2608.14642v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents trained on a single reward signal exploit the gap between the designed reward and the intended behavior. This is part

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

Try Grok 4.6! Grok 4.7, which is a major upgrade, is coming soon.

DGX agent

Try Grok 4.6! Grok 4.7, which is a major upgrade, is coming soon. Grok 4.6 is up to 5× cheaper than GPT-5.6 Sol while delivering the exact same performance on this benchmark Both score 61 on Artificia

model-releaseselon-musk--x
18 Aug 2026
Model Releases

UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations

DGX agent

arXiv:2608.15930v1 Announce Type: new Abstract: Foundation GUI agents can automate complex digital tasks, but deployment is hindered by scarce and biased training data, ambiguous prompts, and unreliab

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Understanding and Stabilizing Deep Q-Learning via Controlled Bootstrapping and Regulated Value Dynamics

DGX agent

arXiv:2608.16182v1 Announce Type: cross Abstract: Deep Q-learning (DQL) has achieved remarkable empirical success in reinforcement learning, yet its training process remains notoriously unstable. Exis

model-releasesarxiv-cs-ai
18 Aug 2026
Research

Unraveling the Size Determination Mechanism of Nanocrystal Synthesis via Interpretable Neural Networks

DGX agent

arXiv:2608.14734v1 Announce Type: cross Abstract: Deep learning models of nanocrystal synthesis enable the prediction of size and shape by encoding precursors and reaction conditions. However, their b

researcharxiv-cs-ai
18 Aug 2026
Model Releases

Unsupervised Anomaly Detection for Image Dataset Quality Assurance in Multi-Center Breast MRI

DGX agent

arXiv:2608.16725v1 Announce Type: cross Abstract: Corrupted, inconsistent, or anomalous data silently threatens the safety and reliability of medical AI. Despite growing regulatory recognition of data

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Valid Per-Field Selective Risk Control for Document Extraction: Three Failure Modes, a Validity Ladder, and When Conditioning Pays

DGX agent

arXiv:2608.14639v1 Announce Type: cross Abstract: Per-field accept/review with selective risk at most alpha -- accept a field only if the error rate among accepted fields is controlled -- is the trust

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Validation-Frontier Representation Selection under Constrained Observation

DGX agent

arXiv:2608.15095v1 Announce Type: new Abstract: AI systems deployed outside clean benchmark settings often rely on observations that are incomplete, unstable, costly, or degraded by monitoring failure

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

Walk Before You Run: The Importance of Data Exploration for Data Analysis Agents

DGX agent

arXiv:2608.16045v1 Announce Type: cross Abstract: LLM-based data-analysis tools are increasingly used to help users analyze messy spreadsheets and workbooks, from answering questions over uploaded fil

model-releasesarxiv-cs-ai
18 Aug 2026
Model Releases

WANDR: A Benchmark for Wide and Deep Research

DGX agent

arXiv:2608.14747v1 Announce Type: new Abstract: WANDR (Wide ANd Deep Research) is a benchmark of 500 realistic, challenging data-collection tasks for research agents. Each task requires a system to di

model-releasesarxiv-cs-lg
18 Aug 2026
Model Releases

We still don’t know how people are really using AI

DGX agent

AI companies like Anthropic and OpenAI regularly publish reports on how people are using products like Claude and ChatGPT, but they only release the data they want us to see, AI researchers say. “Ther

model-releasesmit-tech-review
18 Aug 2026
Model Releases

WeSCE: A Benchmark for Measuring Security Drift in LLM-Driven Code Editing

DGX agent

arXiv:2608.15092v1 Announce Type: cross Abstract: In this work, we introduce WeSCE, a benchmark for quantifying security drift in code editing under weak-security constraints, where tasks specify only

model-releasesarxiv-cs-ai
18 Aug 2026
← Previous
1…776777778779780…1369
Next →