AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
Human
87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,170 results
12 May 2026

Task-Aware Calibration: Provably Optimal Decoding in LLMs

ResearchDGX agent

arXiv:2605.10202v1 Announce Type: cross Abstract: LLM decoding often relies on the model's predictive distribution to generate an output. Consequently, misalignment with respect to the true generating

Task complexity shapes internal representations and robustness in neural networks

ResearchDGX agent

arXiv:2508.05463v2 Announce Type: replace-cross Abstract: Neural networks excel across a wide range of tasks, yet remain black boxes. In particular, how their internal representations are shaped by th

TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder Generation

SafetyDGX agent

arXiv:2605.09810v1 Announce Type: cross Abstract: Protein function is often controlled by ligands that bias the direction of state transitions, such as agonists and antagonists, rather than stabilizin

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Teacher-Aware Evolution of Heuristic Programs from Learned Optimization Policies

Local AiDGX agent

arXiv:2605.10634v1 Announce Type: new Abstract: LLM-based automatic heuristic design has shown promise for generating executable heuristics for combinatorial optimization, but existing methods mainly

Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning

Model ReleasesDGX agent

arXiv:2605.10247v1 Announce Type: new Abstract: Using Large Language Models (LLMs) to process graph-structured data is an active research area, yet current state-of-the-art approaches typically rely o

Teaching Molecular Dynamics to a Non-Autoregressive Ionic Transport Predictor

TutorialsDGX agent

arXiv:2605.09311v1 Announce Type: cross Abstract: Unlike most static material properties widely studied in the machine learning literature, ionic transport properties are inherently dynamic, making th

Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs

SafetyDGX agent

arXiv:2605.09922v1 Announce Type: cross Abstract: While recent self-training approaches have reduced reliance on human-labeled data for aligning LLMs, they still face critical limitations: (i) sensiti

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

Model ReleasesDGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

'tell me about recent funding and news for openai and anthropic' result with @wokeloai mcp included today's trial, tomoro acquisition, and c…

Model ReleasesDGX agent

Yohei Nakajima shared recent funding and news updates about OpenAI and Anthropic on X, including information about a trial involving the wokeloai MCP, a tomorrow acquisition, and additional details (c

TELL-TALE: Task Efficient LLMs with Task Aware Layer Elimination

ResearchDGX agent

arXiv:2510.22767v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) typically come with a fixed architecture, despite growing evidence that not all layers contribute equally to ever

Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding

AgentsDGX agent

arXiv:2507.11198v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) enable new possibilities for qualitative research at scale, including annotation and qualitative coding of educat

Temporal-Decay Shapley: A Time-Aware Data Valuation Framework for Time-Series Data

ResearchDGX agent

arXiv:2605.08153v1 Announce Type: new Abstract: With the rapid development of machine learning applications on time-series data, accurately assessing the value of training samples has become essential

Temporal Sampling Frequency Matters: A Capacity-Aware Study of End-to-End Driving Trajectory Prediction

AgentsDGX agent

arXiv:2605.10388v1 Announce Type: new Abstract: End to end (E2E) autonomous driving trajectory prediction is often trained with camera frames sampled at the highest available temporal frequency, assum

Temporal Tokenization Strategies for Event Sequence Modeling with Large Language Models

ApplicationsDGX agent

arXiv:2512.13618v3 Announce Type: replace Abstract: Representing continuous time is a critical and under-explored challenge in modeling temporal event sequences with large language models (LLMs). Vari

Tensor Product Representation Probes Reveal Shared Structure Across Linear Directions

ResearchDGX agent

arXiv:2605.09967v1 Announce Type: new Abstract: While researchers are finding concepts represented as linear directions in language models, a bag of linear directions fails to capture relational struc

Terminal Matters: Kinodynamic Planning with a Terminal Cost and Learned Uncertainty in Belief State-Cost Space

TutorialsDGX agent

arXiv:2605.09046v1 Announce Type: new Abstract: In many real-world robotic tasks, robots must generate dynamically feasible motions that reliably reach desired goals even under uncertainty. Yet existi

Test-Time Speculation

Model ReleasesDGX agent

arXiv:2605.09329v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by using a fast draft model to generate tokens and a more accurate target model to verify them. Its perfo

Test-Time Training for Visual Foresight Vision-Language-Action Models

ResearchDGX agent

arXiv:2605.08215v1 Announce Type: new Abstract: Visual Foresight VLA (VF-VLA) has become a prominent architectural choice in the recent VLA due to its impressive performance. Nevertheless, the inheren

Text-Guided Multi-Scale Frequency Representation Adaptation

Model ReleasesDGX agent

arXiv:2605.08181v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning methods introduce a small number of training parameters, enabling pre-trained models to adapt rapidly to new data dist

TextBridgeGNN: Pre-training Graph Neural Network for Cross-Domain Recommendation via Text-Guided Transfer

TutorialsDGX agent

arXiv:2601.02366v2 Announce Type: replace-cross Abstract: Graph-based recommendation has achieved great success in recent years. The classical graph recommendation model utilizes ID embedding to store

TFM-Retouche: A Lightweight Input-Space Adapter for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2605.06047v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs), such as TabPFN-2.6, TabICLv2, ConTextTab, Mitra, LimiX, and TabDPT, achieve strong zero-shot performance thr

Thanks to everyone who showcased their projects in today's Hermes Agent Jam!

AgentsDGX agent

Nous Research held a 'Hermes Agent Jam' event where developers showcased projects built using or related to their Hermes agent framework. The event appears to have been a community-driven showcase hig

Thanks to the @huggingface team for adding Hermes Agent to local apps and shipping a native Hermes traces viewer!

Local AiDGX agent

Thanks to the @huggingface team for adding Hermes Agent to local apps and shipping a native Hermes traces viewer! 🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally

The 9 biggest new features in Android 17

IndustryDGX agent

Would it shock you to hear that Android 17 is filled with new AI-enabled features, like improved dictation and vibe-coded widgets? Fortunately, that's not all. The platform is getting non-AI updates t

The Accountability Paradox: How Platform API Restrictions Undermine AI Transparency Mandates

SafetyDGX agent

arXiv:2505.11577v4 Announce Type: replace-cross Abstract: Recent application programming interface (API) restrictions on major social media platforms challenge compliance with the EU Digital Services

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn…

AgentsDGX agent

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn temporal structure in music. Finding temporal structure in

The Agent Use of Agent Beings: Agent Cybernetics Is the Missing Science of Foundation Agents

AgentsDGX agent

arXiv:2605.10754v1 Announce Type: new Abstract: LLM-based foundation agents that perceive, reason, and act across thousands of reasoning steps are rapidly becoming the dominant paradigm for deploying

The Alpha Blending Hypothesis: Compositing Shortcut in Deepfake Detection

Model ReleasesDGX agent

arXiv:2605.10334v1 Announce Type: new Abstract: Recent deepfake detection methods demonstrate improved cross-dataset generalization, yet the underlying mechanisms remain underexplored. We introduce th

The Art of the Jailbreak: Formulating Jailbreak Attacks for LLM Security Beyond Binary Scoring

SafetyDGX agent

arXiv:2605.09225v1 Announce Type: cross Abstract: Jailbreak attacks -- adversarial prompts that bypass LLM alignment through purely linguistic manipulation -- pose a growing operational security threa

The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care

ResearchDGX agent

arXiv:2605.09838v1 Announce Type: new Abstract: Sentiment analysis has been of long-standing interest in psychotherapy research. Recently, the Transformer deep learning architecture has produced text-

The Astonishing Ability of Large Language Models to Parse Jabberwockified Language

ResearchDGX agent

arXiv:2602.23928v2 Announce Type: replace Abstract: We show that large language models (LLMs) have an astonishing ability to recover meaning from severely degraded English texts. Texts in which conten

The Attacker in the Mirror: Breaking Self-Consistency in Safety via Anchored Bipolicy Self-Play

Model ReleasesDGX agent

arXiv:2605.08427v1 Announce Type: new Abstract: Self-play red team is an established approach to improving AI safety in which different instances of the same model play attacker and defender roles in

The autoPET3 Challenge: Automated Lesion Segmentation in Whole-Body PET/CT nicode{x2013} Multitracer Multicenter Generalization

Model ReleasesDGX agent

arXiv:2605.05775v2 Announce Type: replace-cross Abstract: We report the design and results of the third autoPET challenge (MICCAI 2024), which benchmarked automated lesion segmentation in whole-body P

The benchmarks show the gap. NVLS all-reduce latency drops from 586.1µs on H200 to 313.3µs on GB200. In MoE prefill at EP=4, combine falls f…

HardwareDGX agent

The benchmarks show the gap. NVLS all-reduce latency drops from 586.1µs on H200 to 313.3µs on GB200. In MoE prefill at EP=4, combine falls from 730.1µs to 438.5µs. For decode, GB200 sustains much high

The Benefits of Temporal Correlations: SGD Learns k-Juntas from Random Walks Efficiently

ResearchDGX agent

arXiv:2605.10237v1 Announce Type: new Abstract: We study how temporal correlations in the data can make certain sparse learning problems efficiently learnable by gradient-based methods. Our focus is o

The Bystander Effect in Multi-Agent Reasoning: Quantifying Cognitive Loafing in Collaborative Interactions

SafetyDGX agent

arXiv:2605.10698v1 Announce Type: cross Abstract: Multi-agent systems (MAS) assume that collaborating inherently improves Large Language Model (LLM) reasoning. We challenge this by demonstrating that

The Cancellation Hypothesis in Critic-Free RL: From Outcome Rewards to Token Credits

ResearchDGX agent

arXiv:2605.08666v1 Announce Type: new Abstract: A commonly accepted explanation of critic-free RL for LLMs, based on sequence-level rewards, is that it reinforces successful rollouts with a positive a

The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space

ResearchDGX agent

arXiv:2605.09883v1 Announce Type: cross Abstract: As current Multimodal Large Language Models rapidly saturate canonical visual reasoning benchmarks, a key question emerges: do these strong scores gen

The Convergence of Open Table Formats and Open Catalogs: Catalog Commits is Generally Available

IndustryDGX agent

Databricks announced the general availability of Catalog Commits, a feature that integrates open table formats with open catalog systems to enable versioning and tracking of catalog-level changes. Thi

The Differences Between Direct Alignment Algorithms are a Blur

Model ReleasesDGX agent

arXiv:2502.01237v3 Announce Type: replace Abstract: Direct Alignment Algorithms (DAAs) simplify LLM alignment by directly optimizing policies, bypassing reward modeling and RL. While DAAs differ in th

The Direct Integration Theorem: A Rigorous Framework for Consistent Discrete Solutions of the Inverse Radon Problem

ResearchDGX agent

arXiv:2605.09020v1 Announce Type: new Abstract: This paper presents a novel Direct Integration Theorem (DIT), derived as a non-trivial corollary of the classical Central Slice Theorem (CST). The DIT p

The Download: a Nobel winner on AI, and the case for fixing everything

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Three things in AI to watch, according to a Nobel-winning econ

The DSA's Blind Spot: Algorithmic Audit of Advertising and Minor Profiling on TikTok

ResearchDGX agent

arXiv:2603.05653v2 Announce Type: replace-cross Abstract: Adolescents spend an increasing amount of their time in digital environments where their still-developing cognitive capacities leave them unab

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

Model ReleasesDGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

The EDA Primer: From RTL to Silicon

HardwareDGX agent

The EDA Primer covers the semiconductor design and manufacturing workflow, explaining how Electronic Design Automation tools transform Register Transfer Level (RTL) code into physical silicon through

The Ensemble Schr{odinger Bridge filter for Nonlinear Data Assimilation

ResearchDGX agent

arXiv:2512.18928v3 Announce Type: replace Abstract: This work introduces a novel nonlinear optimal filtering method, termed the Ensemble Schr{odinger Bridge nonlinear filter. The proposed filter combi

The Extrapolation Cliff in On-Policy Distillation of Near-Deterministic Structured Outputs

Model ReleasesDGX agent

arXiv:2605.08737v1 Announce Type: cross Abstract: On-policy distillation (OPD) is widely used for LLM post-training. When pushed with a reward-extrapolation coefficient lambda > 1, the student can lif

The finite expression method for turbulent dynamics with high-order moment recovery

TutorialsDGX agent

arXiv:2605.10687v1 Announce Type: new Abstract: Turbulent dynamical systems are characterized by nonlinear interactions and stochastic effects that generate coupled statistical quantities, such as non

The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context Reasoning

AgentsDGX agent

arXiv:2605.10828v1 Announce Type: new Abstract: As large language models are increasingly deployed in retrieval-augmented generation and agentic systems that accumulate extensive context, understandin

The FT says that Amazon employees are doing random unnecessary task automations to consume tokens and to show their bosses that they're usin…

SafetyDGX agent

The FT says that Amazon employees are doing random unnecessary task automations to consume tokens and to show their bosses that they're using AI more https://www.ft.com/content/8ee0d3ef-9548-422d-8ff1

The Generalized Turing Test: A Foundation for Comparing Intelligence

ResearchDGX agent

arXiv:2605.10851v1 Announce Type: new Abstract: We introduce the Generalized Turing Test (GTT), a formal framework for comparing the capabilities of arbitrary agents via indistinguishability. For agen

The Geometric Reasoner: Manifold-Informed Latent Foresight Search for Long-Context Reasoning

ResearchDGX agent

arXiv:2601.18832v3 Announce Type: replace-cross Abstract: Scaling test-time compute enhances long chain-of-thought (CoT) reasoning, yet existing approaches face a fundamental trade-off between computa

The Geometric Structure of Models Learning Sparse Data

SafetyDGX agent

arXiv:2605.08464v1 Announce Type: new Abstract: The manifold hypothesis (MH) is often used to explain how machine learning can overcome the curse of dimensionality. However, the MH is only applicable

The Geometric Wall: Manifold Structure Predicts Layerwise Sparse Autoencoder Scaling Laws

Model ReleasesDGX agent

arXiv:2605.09887v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) operationalise the linear representation hypothesis: they reconstruct model activations as sparse linear combinations of in

The Geometry of Forgetting: Temporal Knowledge Drift as an Independent Axis in LLM Representations

Model ReleasesDGX agent

arXiv:2605.09195v1 Announce Type: new Abstract: Large language models confidently produce outdated answers, and no existing method can detect them. We show this is not an engineering failure but a str

The Global Empirical NTK: Self-Referential Bias and Dimensionality of Gradient Descent Learning

Model ReleasesDGX agent

arXiv:2605.08746v1 Announce Type: new Abstract: In training a neural network with gradient descent (GD), each iteration induces a linear operator that governs first-order updates to a model's internal

The Gordian Knot for VLMs: Diagrammatic Knot Reasoning as a Hard Benchmark

Model ReleasesDGX agent

arXiv:2605.09900v1 Announce Type: new Abstract: A vision-language model can look at a knot diagram and report what it sees, yet fail to act on that structure. KnotBench pairs an 858,318-image corpus f

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans

SafetyDGX agent

arXiv:2605.08837v1 Announce Type: cross Abstract: Abstract concepts - justice, theory, availability - have no single perceivable referent; in the human brain, their meaning emerges from a web of exper

The Impact of Editorial Intervention on Detecting Native Language Traces

ResearchDGX agent

arXiv:2605.10216v1 Announce Type: new Abstract: Native Language Identification (NLI) is the task of determining an author's native language (L1) from their non-native writings. With the advent of huma

The Invisible Handshake: Persistent Overpricing by Adaptive Market Agents

ResearchDGX agent

arXiv:2510.15995v3 Announce Type: replace-cross Abstract: We study overpricing in a repeated game between two representative agents: a market maker, who controls market liquidity, and a market taker,

← Previous
1…10231024102510261027…1453
Next →