AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
Human
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,913 results
11 Jun 2026

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

Model ReleasesDGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

AgentsDGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

Agents All the Way Down; A Methodology for Building Custom AI Agents from Substrate to Production

AgentsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.11869v1 Announce Type: cross Abstract: Custom AI agents areagents that live inside their own application, talk to their own data and tools, enforce their own security boundaries, and carry

Agreement in Representation Space for Open-Ended Self-Consistency

ResearchDGX agent

arXiv:2606.12003v1 Announce Type: new Abstract: Self-consistency improves LLM reasoning by sampling multiple outputs and selecting the most consistent answer, but existing formulations largely rely on

AI Coding Agents Can Reproduce Social Science Findings

Model ReleasesDGX agent

arXiv:2606.11447v1 Announce Type: new Abstract: Recent anecdotal evidence suggests that AI coding agents can reproduce published findings when provided with original data and code; yet systematic eval

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

Model ReleasesDGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

SafetyDGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

AI4Land: Scalable Deep Learning for Global High-Resolution Land Use Reconstruction

HardwareDGX agent

arXiv:2606.11793v1 Announce Type: cross Abstract: Uncertainty in the terrestrial carbon cycle remains a major constraint in climate projections, partly driven by the uncertainties affecting the land s

AI4SLT: Empirical Processes in Lean 4 for Formal Statistical Learning Theory

ResearchDGX agent

arXiv:2602.02285v2 Announce Type: replace-cross Abstract: We present the first comprehensive Lean 4 formalization of statistical learning theory (SLT) grounded in empirical process theory. Our en-to-e

[AINews] Open Models, Model Labs vs Agent Labs, and What's Untrainable — Sarah Guo

AgentsDGX agent

This episode discusses the growing landscape of open-source AI models, contrasts the approaches and philosophies of Model Labs versus Agent Labs in AI development, and explores fundamental limitations

ALIGNBEAM : Inference-Time Alignment Transfer via Cross-Vocabulary Logit Mixing

SafetyDGX agent

arXiv:2606.12342v1 Announce Type: cross Abstract: Domain fine-tuning degrades the safety of large language models: fine-tuned specialists readily comply with harmful prompts framed in domain language.

Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics

Local AiDGX agent

arXiv:2606.12365v1 Announce Type: cross Abstract: We propose Ambient Diffusion Policy, a simple and principled method for imitation learning from suboptimal data in robotics. High-quality, task-specif

An Electric Potential-Augmented Benchmark Dataset for Physics-Guided Image Reconstruction of Electrical Capacitance Tomography

Model ReleasesDGX agent

arXiv:2606.12226v1 Announce Type: new Abstract: While deep learning has significantly advanced image reconstruction of Electrical Capacitance Tomography (ECT), most data-driven methods map directly be

An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations

SafetyDGX agent

arXiv:2606.11218v1 Announce Type: cross Abstract: Ethical deliberation is often misunderstood as a search for single right or wrong answers, creating difficulties for non-ethically trained personnel w

An Ontology-Guided Multi-Anchor Graph Retrieval Framework for Traffic Legal Liability Determination

Model ReleasesDGX agent

arXiv:2606.11910v1 Announce Type: new Abstract: Traffic law liability determination is critical for assigning legal penalties, requiring the simultaneous identification of interdependent statutory pro

An XAI View on Explainable ASP: Methods, Systems, and Perspectives

ResearchDGX agent

arXiv:2601.14764v2 Announce Type: replace Abstract: Answer Set Programming (ASP) is a popular declarative reasoning and problem solving approach in symbolic AI. Its rule-based formalism makes it inher

Analysis: Conservative and Reform UK MPs have increased their volume of posts on X since Elon Musk's takeover, while center-left politicians have reduced theirs (Financial Times)

IndustryDGX agent

Financial Times: Analysis: Conservative and Reform UK MPs have increased their volume of posts on X since Elon Musk's takeover, while center-left politicians have reduced theirs — Majority of Labour a

Analytic Bijections for Smooth and Interpretable Normalizing Flows

ResearchDGX agent

arXiv:2601.10774v2 Announce Type: replace Abstract: A key challenge in normalizing flows is finding expressive invertible scalar bijections. Existing approaches face trade-offs: affine transformations

Anatomically Conditioned Recurrent Refinement for Topology-Aware Circle of Willis Segmentation

ResearchDGX agent

arXiv:2606.12319v1 Announce Type: new Abstract: Segmenting the Circle of Willis (CoW) from Magnetic Resonance Angiography (MRA) is challenging due to complex topology and thin vascular structures that

Anatomy of Post-Training: Using Interpretability to Characterize Data and Shape the Learning Signal

TutorialsDGX agent

arXiv:2606.12360v1 Announce Type: new Abstract: Language-model post-training is the main stage at which model behavior is shaped, yet it still largely involves optimization of scalar rewards that summ

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

Model ReleasesDGX agent

arXiv:2606.11751v1 Announce Type: cross Abstract: Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successi

and now the possible of drastic price cuts, a further sign of OpenAI’s desperation, as foreseen by me 2.5 years ago:

SafetyDGX agent

and now the possible of drastic price cuts, a further sign of OpenAI’s desperation, as foreseen by me 2.5 years ago: Called the LLM price wars – which are about to heat up — over two years ago. See es

Annealed Entropic Allocation for Ranking and Selection

Model ReleasesDGX agent

arXiv:2606.11347v1 Announce Type: cross Abstract: We propose Annealed Entropic Allocation, an annealed weighted soft-min framework for sequential budget allocation in ranking and selection. The centra

Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude

Model ReleasesDGX agent

Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude Big scoop for Maxwell Zeff at Wired: “We’re changing Fable 5’s safeguards for frontier LLM development to make them

Anthropic’s Dario Amodei wants governments to have the power to block ‘dangerous’ AI systems

IndustryDGX agent

Anthropic PBC Chief Executive Dario Amodei is calling on the U.S. government to block the deployment of dangerous artificial intelligence models in the same way as it prevents unsafe airplanes from ta

APEX: A Network-Native Time-Series Foundation Model for Forecasting and Anomaly Detection for Wireless Edge Operations

Model ReleasesDGX agent

arXiv:2606.11553v1 Announce Type: new Abstract: Generic time-series foundation models transfer poorly to wireless network telemetry whose signals are bursty, zero-inflated, and coupled across protocol

APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection

Model ReleasesDGX agent

arXiv:2606.11459v1 Announce Type: cross Abstract: Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While ev

Applied Materials opens a $500M chip equipment manufacturing campus in Singapore, making the city-state home to ~50% of its production capacity alongside the US (Cheng Ting-Fang/Nikkei Asia)

ApplicationsDGX agent

Cheng Ting-Fang / Nikkei Asia: Applied Materials opens a 500M chip equipment manufacturing campus in Singapore, making the city-state home to ~50% of its production capacity alongside the US — SINGAPO

APPO: Agentic Procedural Policy Optimization

SafetyDGX agent

arXiv:2606.12384v1 Announce Type: cross Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents

APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies

SafetyDGX agent

arXiv:2606.12366v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models that couple pretrained Vision-Language Models (VLMs) with continuous action experts have achieved strong manipulatio

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

SafetyDGX agent

arXiv:2606.11634v1 Announce Type: new Abstract: The rapid progress of reasoning and agentic large language models (LLMs) has increased the demand for long-context inference, but self-attention (SA) sc

Are LLMs Bad at Moral Reasoning?

ResearchDGX agent

arXiv:2606.11635v1 Announce Type: cross Abstract: For highly capable AI systems to operate safely in dynamic, open-ended environments, they must be able to identify, understand, and respond to moral r

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation

Model ReleasesDGX agent

arXiv:2606.11670v1 Announce Type: cross Abstract: Subject-preserving video generation is not solved by frontal-face similarity alone: a generated person must remain recognizable across motion, large v

Artificial Intelligence in Ship Finance: Applications, Opportunities, and a Case Study in AI-Augmented Loan Origination

AgentsDGX agent

arXiv:2606.11238v1 Announce Type: cross Abstract: Ship finance is a data-intensive and document-heavy segment of asset-based lending, requiring the integration of financial, technical, contractual, an

As featured today by @FT’s Alphaville: “An entire industry is being propped up by math that is insane.” Available only at Marcus on AI, a “f…

SafetyDGX agent

Gary Marcus critiques fundamental mathematical or computational problems underlying an entire industry, suggesting widespread reliance on flawed or unsound methodologies. The commentary appears on Mar

AsFT: Anchoring Safety During LLM Fine-Tuning Within Narrow Safety Basin

Model ReleasesDGX agent

arXiv:2506.08473v4 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) improves performance but introduces critical safety vulnerabilities: even minimal harmful data can severely

ATLAS: Active Theory Learning for Automated Science

AgentsDGX agent

arXiv:2606.12386v1 Announce Type: cross Abstract: Advancing scientific understanding through mechanistic modeling requires posing the right experimental questions to yield maximally informative data.

Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy

Model ReleasesDGX agent

arXiv:2606.12346v1 Announce Type: cross Abstract: Hematoxylin and eosin (H&E) staining is the cornerstone of histopathology, yet scalable, quantitative analysis of H&E whole-slide images (WSIs) remain

Attention by Synchronization in Coupled Oscillator Networks

ResearchDGX agent

arXiv:2606.12059v1 Announce Type: new Abstract: We address transformer attention on energy-constrained physical substrates. Softmax attention requires exponentiation and global reduction, operations w

Auditing Demographic Bias in Facial Landmark Detection for Fair Human-Robot Interaction

SafetyDGX agent

arXiv:2604.06961v2 Announce Type: replace Abstract: Fairness in human-robot interaction critically depends on the reliability of the perceptual models that enable robots to interpret human behavior. W

Augmenting Molecular Language Models with Local n-gram Memory

SafetyDGX agent

arXiv:2606.12113v1 Announce Type: cross Abstract: Transformer-based language models for SMILES strings suffer from a locality gap: standard character-level tokenization fragments chemically meaningful

Automated Creativity Evaluation of Language Models Across Open-Ended Tasks

AgentsDGX agent

arXiv:2606.11762v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in language understanding, reasoning, and generation, sparking growing interest in thei

Automated Mediator for Human Negotiation: Pre-Mediation via a Structured LLM Pipeline

AgentsDGX agent

arXiv:2606.11379v1 Announce Type: new Abstract: Pre-mediation, the preparatory phase preceding direct human negotiation, plays a critical role in achieving mutually beneficial agreements, yet is often

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

SafetyDGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

AutoMine Solution for AV2 2026 Scenario Mining Challenge

SafetyDGX agent

arXiv:2606.11874v1 Announce Type: new Abstract: With the development of autonomous driving systems, mining high-value, safety-critical, and planning-relevant scenarios from large-scale driving logs ha

Autoregressive Direct Preference Optimization

ResearchDGX agent

arXiv:2602.09533v2 Announce Type: replace Abstract: Direct preference optimization (DPO) has emerged as a promising approach for aligning large language models (LLMs) with human preferences. However,

AVIS: Adaptive Test-Time Scaling for Vision-Language Models

SafetyDGX agent

arXiv:2606.11576v1 Announce Type: cross Abstract: Modern Vision-Language Models (VLMs) benefit from chain-of-thought prompting and test-time scaling, but these gains often come with prohibitive infere

Bad things apparently come in 4’s; tweet below must be updated with the breaking @wsj news that OpenAI is contemplating drastic price cuts, …

SafetyDGX agent

Bad things apparently come in 4’s; tweet below must be updated with the breaking @wsj news that OpenAI is contemplating drastic price cuts, which is surely a sign of weakness. So far today • Banks to

Based Grok 🤣🤣 https://x.com/i/grok/share/32212cc499ae467ebb1f8db2b77d314a

IndustryDGX agent

This entry references Grok, xAI's conversational AI assistant, shared through X's integrated Grok feature. The post appears to be a humorous or casual share by Elon Musk, likely demonstrating Grok's c

Battery detection of XRay images using transfer learning

ResearchDGX agent

arXiv:2606.11779v1 Announce Type: new Abstract: The need for detecting and sorting batteries is drastically increasing for many applications. This study proves the potential of transfer learning in pr

Benchmarking Cross-Domain Audio-Visual Deception Detection

Model ReleasesDGX agent

arXiv:2405.06995v4 Announce Type: replace-cross Abstract: Automated deception detection is crucial for assisting humans in accurately assessing truthfulness and identifying deceptive behavior. Convent

Benchmarking Large Language Models for Safety Data Extraction

Model ReleasesDGX agent

arXiv:2606.11204v1 Announce Type: new Abstract: Accurate extraction of structured information from Safety Data Sheets (SDS) remains challenging in industrial safety due to heterogeneous document forma

Bergson: An Open Source Library for Data Attribution

ResearchDGX agent

arXiv:2606.11660v1 Announce Type: new Abstract: Data attribution is a promising field in interpretability that aims to explain model behavior through the influence of its training data, with applicati

Bernstein-Schur Kernels: Random Features by Sketched Modulation and Radial Randomization

ResearchDGX agent

arXiv:2606.11255v1 Announce Type: new Abstract: Bernstein--Schur kernels are products of a finite-feature kernel (one with an explicit finite-dimensional feature map) and a completely monotone shift-i

Beyond Compaction: Structured Context Eviction for Long-Horizon Agents

Model ReleasesDGX agent

arXiv:2606.11213v1 Announce Type: new Abstract: We present Context Window Lifecycle (CWL), a context-management scheme that gives long-horizon LLM agents an effectively unbounded working horizon. As a

Beyond Dark Knowledge: Mixup-Based Distillation for Reliable Predictions

ResearchDGX agent

arXiv:2606.12171v1 Announce Type: new Abstract: Knowledge Distillation (KD) and mixup have proven effective at inducing smoothness in class boundaries; KD captures inherent class relationships in prob

Beyond Fully Random Masking: Attention-Guided Denoising and Optimization for Diffusion Language Models

ResearchDGX agent

arXiv:2606.12273v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer an efficient alternative to autoregressive models through parallel decoding, yet existing post-training me

Beyond representational alignment with brain-guided language models for robust reasoning

SafetyDGX agent

arXiv:2606.11893v1 Announce Type: cross Abstract: The correspondence between large language models (LLMs) and the neural mechanisms underlying human higher-order cognition remains insufficiently chara

Beyond the Golden Teacher: Enhancing Graph Learning through LLM-GNN Co-teaching

TutorialsDGX agent

arXiv:2606.11583v1 Announce Type: new Abstract: Text-attributed graphs (TAGs) underlie real-world applications such as citation networks, social media, and e-commerce. Few-shot graph learning on TAGs

Beyond Third-Person Audits: Situated Interaction Auditing for User-Centered LLM Bias Research

SafetyDGX agent

arXiv:2606.12247v1 Announce Type: cross Abstract: Research on bias in large language models (LLMs) has predominantly focused on third-person audits, which study how models represent or evaluate demogr

← Previous
1…601602603604605…1516
Next →