AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
13 Jul 2026

Securing the AI supply chain on GKE: Introducing k8s-aibom for automated AI BOMs

Model ReleasesDGX agent

How should your security team manage shadow AI? Workloads deployed by developers without formal registration can often evade traditional security scanners, because organizations are reluctant to slow

10 Jul 2026

EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data

SafetyDGX agent

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scene

Enhancing the KidSat Model: Integrating Geographical Encoding and Data Quality Assessment for Childhood Poverty Prediction

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2607.08281v1 Announce Type: new Abstract: Accurate poverty mapping using satellite imagery is often hindered by (i) noisy and sparse survey-derived supervision, (ii) image quality issues such as

Frequency-Domain Multi-Modality Transportation Modeling

TutorialsDGX agent

arXiv:2607.08475v1 Announce Type: new Abstract: Multi-modality transportation refers to urban systems composed of multiple transportation modes, such as traffic flow and public transit, whose dynamics

Functional and Secure Code Generation with Task Vectors

Model ReleasesDGX agent

arXiv:2607.07881v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for code generation, but they struggle to generate functional code free of security vulnerabilities

LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity

Model ReleasesDGX agent

arXiv:2607.08152v1 Announce Type: cross Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language

LTM: Large-scale Terrain Model for Wildfire-prone Landscapes

SafetyDGX agent

arXiv:2607.08711v1 Announce Type: new Abstract: Accurate 3D terrain maps are essential for emergency response when assessing wildfire hazards. However, wildfire-prone regions often span vast areas whe

Open-source AI model developer MiniMax raises $2B in funding

IndustryDGX agent

MiniMax Group Inc., a Shanghai-based artificial intelligence developer, is raising 2 billion in funding. Bloomberg reported on Thursday that more than half the funds are expected to come from the sale

PARA-PV: Physics-Aware Retrieval-Augmented PV Prediction Based on Frozen Foundation Model and Distribution Shift Correction

Local AiDGX agent

arXiv:2607.08079v1 Announce Type: new Abstract: Accurate photovoltaic (PV) power forecasting is essential for reliable grid dispatch and renewable energy integration, yet it remains challenging becaus

Reaction-network reasoning with frontier models for experimentally confirmed catalyst-selectivity hypotheses

ApplicationsDGX agent

arXiv:2607.08003v1 Announce Type: cross Abstract: Catalysts are essential for sustainable chemical manufacturing, yet discovering novel architectures remains a bottleneck dominated by trial-and-error

Search-based Testing of Vision Language Models for In-Car Scene Understanding

SafetyDGX agent

arXiv:2607.02300v2 Announce Type: replace Abstract: In the automotive domain, in-car scene understanding (ISU) enables the detection of safety-critical events, such as driver distraction, and supports

Secure Decentralized Federated Learning via Gossip and Virtual Voting

Model ReleasesDGX agent

arXiv:2607.08651v1 Announce Type: new Abstract: Decentralized federated learning (DFL) removes the central server by letting nodes exchange model updates through peer-to-peer gossip, but existing goss

Unpaired Joint Distribution Modeling via Multi-Scale Image Representations

ApplicationsDGX agent

arXiv:2607.08198v1 Announce Type: new Abstract: This paper studies the problem of learning a joint distribution from marginal observations, which is inherently ill-posed due to the ambiguity of feasib

9 Jul 2026

Attention in Geometry: Scalable Spatial Modeling via Adaptive Density Fields and FAISS-Accelerated Kernels

ResearchDGX agent

arXiv:2601.06135v3 Announce Type: replace-cross Abstract: Spatial computation in geographic systems increasingly requires query-conditioned, local, interpretable aggregation under metric constraints.

Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages

Model ReleasesDGX agent

arXiv:2607.06596v1 Announce Type: cross Abstract: Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicio

Generative Diffusion Models of Stochastic Graph Signals

TutorialsDGX agent

arXiv:2607.06833v1 Announce Type: new Abstract: Sampling stochastic signals supported on a graph underlies many graph machine learning tasks, including recommender systems, forecasting in financial ma

Geometric Self-Distillation for Reasoning Generalization

Model ReleasesDGX agent

arXiv:2607.06855v1 Announce Type: cross Abstract: On-policy distillation is a practical post-training recipe for large language models, supplying dense teacher supervision on the student's own traject

Grok 4.5 is also rank 1 in SWE marathon

Model ReleasesDGX agent

Grok 4.5, xAI's large language model, has achieved the top ranking in the SWE (Software Engineering) Marathon benchmark, according to an announcement by Elon Musk. This ranking suggests the model demo

Looks like Grok 4.5 is #1 on at least a few benchmarks. Better than expected.

Model ReleasesDGX agent

Elon Musk announced that Grok 4.5, an AI model developed by xAI, has achieved top performance on several benchmarks, exceeding prior expectations. The post suggests the model's capabilities have surpa

On Adversarial Vulnerability of Vision-Language Models through the Lens of Intermediate Spectral Subspaces

ResearchDGX agent

arXiv:2607.07375v1 Announce Type: cross Abstract: Adversarial vulnerability in deep neural networks (DNNs) has been studied from the perspectives of decision-boundary geometry, feature robustness, inp

The Appeal and Reality of Recycling LoRAs with Adaptive Merging

Model ReleasesDGX agent

arXiv:2602.12323v2 Announce Type: replace Abstract: The widespread availability of fine-tuned LoRA modules for open pre-trained models has led to an interest in methods that can adaptively merge LoRAs

The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI

Model ReleasesDGX agent

arXiv:2607.06906v1 Announce Type: new Abstract: Agentic AI development today runs on token maxing: buying capability with tokens -- longer reasoning traces, more turns, wider tool payloads, bigger rep

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very s…

TutorialsDGX agent

We stand at a critical crossroads in the debate over AI governance in the United States, and it feels like we are inching closer to a very serious battle over whether or not open source models will ev

8 Jul 2026

Estimation of instrument and noise parameters for inverse problem based on prior diffusion model

ResearchDGX agent

arXiv:2602.11711v2 Announce Type: replace-cross Abstract: This article addresses the issue of estimating observation parameters (response and error parameters) in inverse problems. The focus is on cas

EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems

AgentsDGX agent

arXiv:2607.05638v1 Announce Type: cross Abstract: Teams deploying large language models in business contexts need evaluation systems, yet most treat evaluation as static model selection: run benchmark

From Blueprint to Reality: Modeling and Applying Putnam's Social Capital Theory with LLM-based Multi-agent Simulations

SafetyDGX agent

arXiv:2607.06080v1 Announce Type: cross Abstract: Putnam's Social Capital Theory is a foundational framework for collective action and community prosperity. However, traditional empirical methods face

Google Cloud named Leader in the 2026 Gartner® Magic Quadrant™ for AI Infrastructure

Model ReleasesDGX agent

In the agentic era, AI is evolving from answering questions to reasoning and taking action. Companies who want to lead in this next phase of AI need computing infrastructure that’s designed and optimi

i-EXAM: Instructable and Explainable Attack Connectivity Graph Modeler

ResearchDGX agent

arXiv:2607.05888v1 Announce Type: cross Abstract: i-EXAM is a planning-powered tool that helps system administrators to create security profiles of complex networks and perform what-if analyses to ide

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

Model ReleasesDGX agent

arXiv:2607.05458v1 Announce Type: cross Abstract: Large language model (LLM) agents are usually improved by changing prompts, models, or hand-written workflows, while the execution harness around the

LongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction Synthesis

Model ReleasesDGX agent

arXiv:2607.06160v1 Announce Type: cross Abstract: Synthesizing long-context supervised fine-tuning (SFT) data is a scalable way to enhance the long-context understanding of large language models (LLMs

More info about speculative decoding with llama.cpp: https://github.com/ggml-org/llama.cpp/blob/master/docs/speculative.md

Model ReleasesDGX agent

Speculative decoding is a technique implemented in llama.cpp that speeds up inference by using a smaller, faster model to predict multiple tokens ahead, which a larger model then verifies in parallel,

Robust Face Super-Resolution and Recognition Through Multi-Feature Aggregation in Diffusion Models

TutorialsDGX agent

arXiv:2607.05702v1 Announce Type: new Abstract: Images acquired in surveillance environments often suffer from conditions such as low resolution, variations in pose, irregular illumination, and occlus

SecureCode: A Production-Grade Multi-Turn Dataset for Training Security-Aware Code Generation Models

AgentsDGX agent

arXiv:2512.18542v3 Announce Type: replace-cross Abstract: AI coding assistants produce vulnerable code in 45% of security-relevant scenarios~ite{veracode2025}, yet no public training dataset teaches b

Self-Review Reinforcement Learning (SRRL) with Cross-Episode Memory and Policy Distillation

Model ReleasesDGX agent

arXiv:2607.05541v1 Announce Type: cross Abstract: Reinforcement Learning is commonly used to train large language models using environmental feedback. In applied settings, the environment usually prov

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows

Model ReleasesDGX agent

arXiv:2607.06229v1 Announce Type: cross Abstract: Major cloud data platforms now expose large language model capabilities as native SQL functions, enabling analysts to perform classification, filterin

7 Jul 2026

Autonomous Search for Sparsely Distributed Visual Phenomena through Environmental Context Modeling

AgentsDGX agent

arXiv:2603.10174v2 Announce Type: replace Abstract: Autonomous underwater vehicles (AUVs) are increasingly used to survey coral reefs, yet efficiently locating specific coral species of interest remai

Bayesian Invariance Modeling of Multi-Environment Data

ResearchDGX agent

arXiv:2506.22675v4 Announce Type: replace-cross Abstract: Invariant prediction [Peters et al., 2016] analyzes feature/outcome data from multiple environments to identify invariant features - those wit

Beyond Multilingual Averages: MTEB-PT, a Benchmark for Portuguese Sentence Encoders

Model ReleasesDGX agent

arXiv:2607.04071v1 Announce Type: cross Abstract: Portuguese remains underrepresented in text embedding evaluation, despite being one of the most widely spoken languages in the world. As a result, emb

CLEANER: Self-Purified Trajectories Boost Agentic Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.15141v2 Announce Type: replace Abstract: Agentic Reinforcement Learning (RL) has empowered Large Language Models (LLMs) to utilize tools like Python interpreters for complex problem-solving

Conditional Clifford-Steerable CNNs for PDE Modeling

ResearchDGX agent

arXiv:2510.14007v2 Announce Type: replace-cross Abstract: We introduce Conditional Clifford-Steerable CNNs (C-CSCNNs), a unified framework that incorporates equivariance to arbitrary pseudo-Euclidean

Data modeling patterns for Amazon Quick Sight multi-dataset relationships

IndustryDGX agent

In this post, we shift from concepts to patterns. For each schema, you’ll find a table structure, use cases, implementation steps, and sample SQL queries. We also cover workarounds for advanced scenar

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval

Model ReleasesDGX agent

arXiv:2606.20280v2 Announce Type: replace-cross Abstract: Leveraging Multimodal Large Language Models (MLLMs) via contrastive learning has become a mainstream paradigm for improving the performance of

FastCSP: Accelerated Molecular Crystal Structure Prediction with Universal Model for Atoms

ResearchDGX agent

arXiv:2508.02641v2 Announce Type: replace-cross Abstract: Molecular crystal structure prediction (CSP) is essential for applications in pharmaceuticals and organic electronics. However, CSP remains ch

FLOAT Drone for Physical Interaction: Lateral Airflow Reduction, Wrench Modeling, and Adaptive Control

SafetyDGX agent

arXiv:2607.04260v1 Announce Type: new Abstract: Aerial physical interaction represents a promising direction for next-generation unmanned aerial vehicles (UAVs), but it requires an aerial platform tha

From Fixed to Free Cameras: Calibration-Free View-Robust Vision-Language-Action Model

SafetyDGX agent

arXiv:2607.05396v1 Announce Type: cross Abstract: Real-world robot deployment rarely maintains the training-stage camera setup, where cameras often experience repositioning or remounting depending on

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some s…

Model ReleasesDGX agent

I wonder how this could be used to study creative thinking and the generation of new ideas. If I understand correctly, the J-space is some sort of internal workspace where the model holds concepts to

Incentivizing Vision Language Models to Search for Long Video Question Answering

AgentsDGX agent

arXiv:2607.02959v1 Announce Type: new Abstract: We introduce VSeek, an agentic framework that transforms long-video question answering (LVQA) from a passive, single-pass perception task into a multi-t

Integrated Graph Search and Model Predictive Control for Smooth and Efficient Path Planning in Autonomous Vehicles

SafetyDGX agent

arXiv:2607.04259v1 Announce Type: new Abstract: Path planning is a fundamental component of autonomous vehicles, where achieving safe, comfortable, and dynamically feasible paths while ensuring comput

Multi-Large Language Model Orchestrated Severity Assessment of Clinical Records (MOSAIC)

Local AiDGX agent

arXiv:2607.05032v1 Announce Type: new Abstract: Background: Disease severity is a multidimensional construct difficult to capture with rule-based approaches in Electronic Healthcare Records (EHR). Age

Obey, Diverge, Collapse: Blind Obedience to Incorrect Instructions Drives Code LLMs to Irrecoverable Code Semantic Collapse

Model ReleasesDGX agent

arXiv:2607.04537v1 Announce Type: cross Abstract: Code language models are now trusted collaborators in production workflows for debugging, refactoring, and iterative repair, and every benchmark that

Re:Form -- Reducing Human Annotations in Scalable Formal Software Verification with RL in LLMs: A Preliminary Study on Dafny

Model ReleasesDGX agent

arXiv:2507.16331v4 Announce Type: replace Abstract: Existing informal language-based (e.g., human language) Large Language Models (LLMs) trained with Reinforcement Learning (RL) face a significant cha

Robust Counterfactual Explanations under Model Multiplicity Using Multi-Objective Optimization

ResearchDGX agent

arXiv:2501.05795v4 Announce Type: replace-cross Abstract: In recent years, explainability in machine learning has gained importance. In this context, counterfactual explanation (CE), which is an expla

Self-Reference in Large Language Models: The Introspection Threshold for Recursive Self-Improvement

SafetyDGX agent

arXiv:2607.04277v1 Announce Type: cross Abstract: The pursuit of self-evolving AI raises a critical question: when is autonomous self-improvement sustainable rather than degenerative? Drawing an analo

SynSFX: Multi-Model Sound Effects Synthesis Dataset for Deepfake Detection and Evaluation

ResearchDGX agent

arXiv:2607.04848v1 Announce Type: cross Abstract: While audio deepfake detection has advanced significantly, representative detectors show limited generalization to synthetic sound effects. Existing e

TexTailor: Inference-Time Textual Guidance Tailoring for Multimodal Diffusion Transformers

Model ReleasesDGX agent

arXiv:2601.02211v2 Announce Type: replace Abstract: Recent breakthroughs of transformer-based diffusion models, particularly with Multimodal Diffusion Transformers (MMDiT) driven models like FLUX and

The problem with Anthropic's consciousness paper My last post got more attention than I expected, and the question I keep getting is some ve…

Model ReleasesDGX agent

The problem with Anthropic's consciousness paper My last post got more attention than I expected, and the question I keep getting is some version of 'okay, so what is actually wrong with the paper?'.

6 Jul 2026

How Open Models Are Driving AI Research

HardwareDGX agent

Every year, the International Conference on Machine Learning (ICML) reveals where thousands of AI researchers have decided to put their work. This year’s accepted papers reveal a clear direction: open

Must-read research by Anthropic. Here is the simple explanation and why this is a big deal. We suspect LLMs perform 'internal reasoning'. Bu…

Model ReleasesDGX agent

Must-read research by Anthropic. Here is the simple explanation and why this is a big deal. We suspect LLMs perform 'internal reasoning'. But little is known or do good methods exist to understand it.

Teaching models to forget: Selective unlearning with Amazon Nova

IndustryDGX agent

In this post, we introduce Reverse Direct Preference Optimization (rDPO), the novel unlearning technique behind Amazon Nova Customizable Content Moderation Settings (CCMS), and show how it reduces ove

3 Jul 2026

ADMC: Attention-based Diffusion Model for Missing Modalities Feature Completion

ResearchDGX agent

arXiv:2507.05624v2 Announce Type: replace Abstract: Multimodal emotion and intent recognition is essential for automated human-computer interaction, It aims to analyze users' speech, text, and visual

← Previous
1…248249250251252…1018
Next →