AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,316
  • Agents7,549
  • Applications5,409
  • Concepts5
  • Hardware1,833
  • Industry6,164
  • Local Ai4,927
  • Model Releases23,845
  • Research20,123
  • Safety13,368
  • Syntheses17
  • Tools1,675
  • Tutorials3,401

Source
HumanDGX agent

88,316Total entries
1Added by human
88,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,550 results
14 Aug 2026

MergeOver: Post-Training Token Merging for Recursive Vision Transformers

Model ReleasesDGX agent

arXiv:2608.13141v1 Announce Type: new Abstract: Vision Transformers (ViTs) demonstrate exceptional performance in computer vision but suffer from large parameter counts and quadratic computational com

Novel Knowledge-Guided Generative Methods for Synthetic Transcriptomic Data

Model ReleasesDGX agent

arXiv:2608.13256v1 Announce Type: cross Abstract: As biomedical research increasingly relies on data-intensive tools, the quality and utility of datasets are critical. Challenges such as imbalances, b

On Measuring Semantic Preservation in Legal Ontology Learning

ApplicationsDGX agent

arXiv:2608.12326v1 Announce Type: new Abstract: Ontology learning transforms unstructured text into structured representations for automated reasoning. Yet structuring information risks losing it, and

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Online Correlation Clustering: Simultaneously Optimizing All ell_p-norms

SafetyDGX agent

arXiv:2510.15076v2 Announce Type: replace Abstract: The ell_p-norm objectives for correlation clustering present a fundamental trade-off between minimizing total disagreements (the ell_1-norm) and ens

PseudoMapLabeler: Confidence-Aware Pseudo-Label Generation for Semi-Supervised Online Mapping

ApplicationsDGX agent

arXiv:2608.12600v1 Announce Type: cross Abstract: A critical challenge in deploying online HD map construction systems to real-world scenarios is the scarcity of labeled training data, which limits mo

Qwen3.8-Max is live on Fireworks, ready for agents, heavy coding, and long-context work. Fireworks is right there with us as a Day 0 launch …

Model ReleasesDGX agent

Qwen3.8-Max is live on Fireworks, ready for agents, heavy coding, and long-context work. Fireworks is right there with us as a Day 0 launch partner. What a way to kick things off!It's Day 0, cue the F

SpinCastML an Open Decision-Making Application for Inverse Design of Electrospinning Manufacturing: A Machine Learning, Optimal Sampling and Inverse Monte Carlo Approach

Model ReleasesDGX agent

arXiv:2602.09120v2 Announce Type: replace Abstract: Electrospinning is a powerful technique for producing micro to nanoscale fibers with application specific architectures. Small variations in solutio

StrAD: A Streaming Method and Benchmark for Audio Description Generation for Long-form Videos

Model ReleasesDGX agent

arXiv:2608.12549v1 Announce Type: new Abstract: Visual content is the dominant medium of communication, yet without audio descriptions (ADs), it remains inaccessible to blind and low-vision people. AD

The difference between 'medium' and 'xhigh' reasoning effort for Qwen3.8-27B is actually insane.

Model ReleasesDGX agent

I'm currently testing out Qwen3.8-27B using Unsloth's UD-Q4_K_XL running a freshly rebuilt llama.cpp. I have a 22GB RTX 2080TI on which I'm able to fit 100k context with q8_0 quantization, and using M

The Impact of Temporal Context Length and Encoding Strategies on Self-Supervised ECG Representation Learning

ApplicationsDGX agent

arXiv:2608.12695v1 Announce Type: new Abstract: Self-supervised electrocardiogram (ECG) models are often trained on a few seconds of ECG signal and, increasingly, on discretized token sequences. It re

The Role of Natural Language Understanding in Multimodal Video-Based Dengue Diagnosis

SafetyDGX agent

arXiv:2608.12677v1 Announce Type: new Abstract: Detecting infection-related behavioral changes in mosquitoes from video data is challenging because mosquitoes are small, move rapidly and irregularly,

Topology-Unified 2D Pose Estimation across Intact, Residual and Prosthetic Limbs

Model ReleasesDGX agent

arXiv:2608.13047v1 Announce Type: new Abstract: Driven by the availability of large-scale datasets, Human Pose Estimation (HPE) plays a critical role in numerous downstream tasks. However, mainstream

UltraFast-LiNET: Light-weight multi-scale shift convolutional network for real-time low-light image enhancement

Model ReleasesDGX agent

arXiv:2512.02965v2 Announce Type: replace Abstract: Addressing the urgent need for high-performance real-time low-light image enhancement on resource-constrained edge devices in low-illumination scena

v0.32.11

Model ReleasesDGX agent

What's Changed launch: add Muse Code integration by @dhiltgen in #17594 model/renderers: match Muse Glimmer reasoning template by @dhiltgen in #17732 launch: add DeepSeek Harness integration by @Parth

Where Should Optimizer State Live? Tiered State Allocation for Memory-Efficient Mixture-of-Experts Training

Model ReleasesDGX agent

arXiv:2607.19058v2 Announce Type: replace-cross Abstract: Optimizer state is the largest single line item in the memory budget of mixture-of-experts (MoE) training. On a 6.78B-parameter MoE language m

13 Aug 2026

A Cascaded Unsupervised-Supervised NLP Pipeline for Detecting Accusatory Language in Public Procurement

Model ReleasesDGX agent

arXiv:2608.12269v1 Announce Type: new Abstract: Public procurement involves the allocation of substantial financial resources; therefore, continuous oversight through audits, controls, and monitoring

Accuracy and Order Sensitivity Diverge Under Label-Free Strategies

SafetyDGX agent

arXiv:2608.11947v1 Announce Type: cross Abstract: Multiple-choice benchmarks are widely used to evaluate large language models, but MCQ scores conflate knowledge with sensitivity to option order, whic

AI Guardrail Survival under Single-Cycle Agentic Self-Summarization

SafetyDGX agent

arXiv:2608.11392v1 Announce Type: cross Abstract: Long-running agents periodically compact their context, replacing the transcript with a model-generated summary.Recent work shows that dropping a stan

ai is the ultimate grow the pie event cheaper inference = better gross margins for application companies, AND beautiful cohort expansion for…

IndustryDGX agent

ai is the ultimate grow the pie event cheaper inference = better gross margins for application companies, AND beautiful cohort expansion for model and inference companies Sequoia Capital GP @sonyatwee

Anthropic details multiagent experiments showing Claude agents can wage a 'turf war' over incompatible goals, fail to coordinate, collude on prices, and more (Rebecca Bellan/TechCrunch)

Model ReleasesDGX agent

Rebecca Bellan / TechCrunch: Anthropic details multiagent experiments showing Claude agents can wage a “turf war” over incompatible goals, fail to coordinate, collude on prices, and more — What happen

Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence

Model ReleasesDGX agent

arXiv:2608.11341v1 Announce Type: new Abstract: Apollo did not reach the Moon merely because its engineers could solve difficult equations. It succeeded by turning a distant ambition into a mission ar

Automated binary classification of hazelnut X-ray images: A deep-learning benchmark for quality assessment

Model ReleasesDGX agent

arXiv:2608.11759v1 Announce Type: new Abstract: Non-destructive X-ray imaging can reveal internal hazelnut defects that are difficult to detect by external inspection alone; however, automated interpr

Backtrader-Bench: Benchmarking LLM Agents on Algorithmic Trading with Self-Generated MCQs

Model ReleasesDGX agent

arXiv:2608.11232v1 Announce Type: cross Abstract: Evaluating LLM coding agents in algorithmic trading is difficult because static benchmarks risk data contamination and numerical backtest outputs requ

BoltNet: An Ultra-Lightweight Convolutional Network for On-Device Plant Species Identification

Local AiDGX agent

arXiv:2608.11844v1 Announce Type: new Abstract: Automated plant species identification from citizen-science imagery is an established, demanding fine-grained recognition problem: large taxonomic label

Causal inference for group-contaminated structured outcomes: observable quotients, lossless reduction and exact randomization inference

Model ReleasesDGX agent

arXiv:2608.11954v1 Announce Type: cross Abstract: Structured potential outcomes such as microscopy images may be recorded after an unknown, unit-specific transformation. If that transformation can dep

CoDiR: Confidence-Guided Diffusion Refinement for Semi-Supervised Histopathology Segmentation

Model ReleasesDGX agent

arXiv:2608.11807v1 Announce Type: new Abstract: Semi-supervised histopathology segmentation is challenging due to scarce annotations and unreliable pseudo-labels in ambiguous gland regions. To address

Confucius4-TTS: Transcript-Free Cross-Lingual Zero-Shot TTS with a Learnable Speaker Encoder

Model ReleasesDGX agent

arXiv:2608.11650v1 Announce Type: cross Abstract: Recent advances in zero-shot text-to-speech (TTS) have substantially improved speech quality and voice cloning fidelity. However, many zero-shot TTS s

Convergence Guarantees of Gradient Descent for Neural Networks via Generalized Lipschitz Smoothness

Model ReleasesDGX agent

arXiv:2608.11479v1 Announce Type: new Abstract: We establish convergence guarantees of gradient descent for general feedforward neural networks of arbitrary width or depth, with no special requirement

CORE-3D: Context-aware Open-vocabulary Retrieval by Embeddings in 3D

Model ReleasesDGX agent

arXiv:2509.24528v4 Announce Type: replace-cross Abstract: Object retrieval from a scene has become a new trend of research due to its numerous applications. Recent approaches achieve zero-shot, open-v

CTBench: Evaluating Troubleshooting Capabilities of AI Agents in Realistic Telecom Network Operations

Model ReleasesDGX agent

arXiv:2608.12002v1 Announce Type: new Abstract: Agents are increasingly considered for automating network operations and maintenance, where engineers must diagnose network faults, optimize configurati

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and …

Model ReleasesDGX agent

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Co

DreamFly: Causal Memory and Receding-Horizon Diffusion Planning for Aerial Vision-Language Navigation

Model ReleasesDGX agent

arXiv:2608.12308v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires an embodied agent to integrate visual evidence over time, plan future actions, and determine when it

Easper: An Accessible ASR Pipeline for Language Documentation

ResearchDGX agent

arXiv:2608.11629v1 Announce Type: new Abstract: Audio transcription is a critical bottleneck in language documentation. While multilingual Automatic Speech Recognition (ASR) models like Whisper offer

Federated Learning for the Design of Parametric Insurance Indices under Heterogeneous Renewable Production Losses

Model ReleasesDGX agent

arXiv:2601.12178v2 Announce Type: replace Abstract: We propose a federated learning framework for the calibration of parametric insurance indices under heterogeneous renewable energy production losses

Glance, Scrutinize, and Think: Advancing Video Anomaly Detection from Training-Free to Agentic Reasoning

Model ReleasesDGX agent

arXiv:2608.11260v1 Announce Type: new Abstract: Video Anomaly Detection (VAD) aims to identify anomalous events and localize their temporal intervals. Existing approaches exhibit a 'when-what' dissoci

Grok 4.6 ranks #1 on the GPQA Diamond leaderboard 🧠 Grok 4.6 (high) scores 95% - the highest score on the chart for graduate-level scientif…

Model ReleasesDGX agent

Grok 4.6 ranks #1 on the GPQA Diamond leaderboard 🧠 Grok 4.6 (high) scores 95% - the highest score on the chart for graduate-level scientific reasoning It outperforms Claude Fable 5, Opus 5, GPT-5.6 S

Group Alignment-Induced Sycophancy: A Two-Sided Evaluation of Steerable Pluralistic Alignment

SafetyDGX agent

arXiv:2608.11528v1 Announce Type: new Abstract: Group alignment adapts a language model to a demographic group to produce responses that reflect the group's opinions, values, and preferences. Sycophan

Hand Visibility Detector: Per-Keypoint Visibility Estimation for Hands

Model ReleasesDGX agent

arXiv:2608.11574v1 Announce Type: new Abstract: Hand Pose Estimation (HPE) is a fundamental technology for various applications such as AR/VR and robotics. In these applications, the visibility of eac

HandEdit: A Unified Benchmark for Egocentric Human-to-Robot Dexterous Hand Image Editing

Model ReleasesDGX agent

arXiv:2608.12122v1 Announce Type: cross Abstract: Robotic manipulation with dexterous hands is a cornerstone of Embodied AI, yet its progress is stifled by the high cost of collecting embodiment-aware

Hybrid-Policy Self-Editing for Composable Unstructured Knowledge Editing

SafetyDGX agent

arXiv:2608.11660v1 Announce Type: cross Abstract: Large language models (LLMs) achieve remarkable performance across natural language tasks, yet they are trained on static corpora and their knowledge

I asked DeepSeek-V4-Flash to work with Muse-Glimmer for Vision ability in PI agent and it produced this

Model ReleasesDGX agent

Same old prompt, just appended a TIP in the end: 'Write a single HTML file with a full-page canvas and no libraries. Simulate a realistic side-view of a moving car as the main subject. Keep the car vi

Localizing to Debias: A Patch-Level Benchmark and Baseline for Weakly Supervised Spatial Anomaly Detection

Model ReleasesDGX agent

arXiv:2608.12045v1 Announce Type: new Abstract: Despite growing interest in weakly supervised video anomaly detection (WSVAD), current methods struggle to bridge the gap between coarse temporal superv

Long-Horizon Forecasting of Complete Financial Statements with Forma

Model ReleasesDGX agent

arXiv:2608.11327v1 Announce Type: new Abstract: Specialist training beats generalist scale when forecasting financial statements. To our knowledge, no prior work jointly forecasts complete financial s

MicroAUNet: Boundary-Enhanced Multi-scale Fusion with Knowledge Distillation for Colonoscopy Polyp Image Segmentation

Model ReleasesDGX agent

arXiv:2511.01143v2 Announce Type: replace-cross Abstract: Early and accurate segmentation of colorectal polyps is critical for reducing colorectal cancer mortality, which has been extensively explored

My cofounder Tim, whom you will not find on X, shares some of the strategic thinking behind our latest updates https://venturebeat.com/infra…

Model ReleasesDGX agent

My cofounder Tim, whom you will not find on X, shares some of the strategic thinking behind our latest updates https://venturebeat.com/infrastructure/mistral-ai-wants-to-build-1-gigawatt-of-european-c

Ollama Pro annual subscription feels borderline scam at this point: $200 paid, Kimi K3 paywalled, lower limits, and still no refund response

Local AiDGX agent

I’m genuinely pretty disappointed with how my Ollama Pro annual subscription turned out. For context, I’m a technology editor and developer, and I had actually used Ollama Pro before. My previous expe

On Benchmarking Human-Like Intelligence in Machines

ResearchDGX agent

arXiv:2502.20502v2 Announce Type: replace Abstract: Recent advances in Artificial Intelligence (AI) have yielded powerful computational models that, by learning from vast amounts of human-generated da

OpenAI previews Ultrafast, an API tier powered by Cerebras that runs GPT-5.6 Sol up to 14× faster and generates up to 750 output tokens per second (Zac Hall/9to5Mac)

Model ReleasesDGX agent

Zac Hall / 9to5Mac: OpenAI previews Ultrafast, an API tier powered by Cerebras that runs GPT-5.6 Sol up to 14× faster and generates up to 750 output tokens per second — OpenAI is previewing a new way

OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing

Model ReleasesDGX agent

arXiv:2512.07826v3 Announce Type: replace Abstract: The quality and diversity of instruction-based image editing datasets are continuously increasing, yet large-scale, high-quality datasets for instru

RealisticTritonBench: A Benchmark for Triton-Kernel Generation in Real-World AI Frameworks

Model ReleasesDGX agent

arXiv:2608.12004v1 Announce Type: cross Abstract: In modern AI frameworks, GPU kernels are key to overall system performance. Combining usability, portability, and near-handwritten CUDA performance, T

Regime-Gated Residual Mixture-of-Experts for Cross-Sectional Volatility Forecasting

ResearchDGX agent

arXiv:2608.12251v1 Announce Type: cross Abstract: Financial volatility is regime dependent, yet incorporating regime information into neural networks can also destabilize training. This paper asks whe

SCAR-GS: Spatial Context Attention for Residuals in Progressive Gaussian Splatting

ResearchDGX agent

arXiv:2601.04348v2 Announce Type: replace Abstract: Recent advances in 3D Gaussian Splatting have allowed for real-time, high-fidelity novel view synthesis. Nonetheless, these models have significant

SteeringSafety: Benchmarking Representation Steering in LLMs Across Safety Perspectives

Model ReleasesDGX agent

arXiv:2509.13450v3 Announce Type: replace Abstract: We introduce SteeringSafety, a benchmark for evaluating representation steering methods across nine safety perspectives spanning 18 datasets. While

The Next Challenge for Agentic Cybersecurity: A Realistic, Contamination-Free Reverse Engineering Benchmark

Model ReleasesDGX agent

arXiv:2608.11469v1 Announce Type: cross Abstract: AI agents are rapidly improving in cybersecurity capabilities when the source code is available for analysis, yet much of the software most consequent

There is much less signal in agent leaderboards than the rankings imply. A four-facet Generalizability Theory decomposition across TheAgentC…

AgentsDGX agent

There is much less signal in agent leaderboards than the rankings imply. A four-facet Generalizability Theory decomposition across TheAgentCompany, tau-squared-bench, and AppWorld finds the agent main

Trained a 1.5B to write shell commands so I'd stop googling tar flags. Runs on a laptop CPU in ~1 sec.

Model ReleasesDGX agent

I've been googling 'tar extract gz' for about ten years. and I finally did something about it. It started out as a research project and I ended up with a Fine-tuned Qwen2.5-Coder-1.5B on 125k natural-

You don't stand up an agent. You upload a folder. Claude Code is a harness for your laptop. Managed Deep Agents is a harness for production.…

Model ReleasesDGX agent

You don't stand up an agent. You upload a folder. Claude Code is a harness for your laptop. Managed Deep Agents is a harness for production. Want Slack? Add a file. Want a daily run? Add a file. Want

12 Aug 2026

A Systematic Sample Size Analysis of ML-Based Path Loss Prediction for LPWAN

ApplicationsDGX agent

arXiv:2608.11083v1 Announce Type: cross Abstract: Low Power Wide Area Networks like LoRa are increasingly deployed for smart city applications, requiring accurate path loss prediction for effective ne

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses

Model ReleasesDGX agent

arXiv:2606.08348v2 Announce Type: replace Abstract: LLM agents increasingly rely on prompts, tools, memory, SOPs, skills, and harness feedback, yet current self-evolution pipelines often update these

Beyond a Bag of Features: Set-Level Instability in Sparse Autoencoders

ResearchDGX agent

arXiv:2608.11197v1 Announce Type: cross Abstract: Shani et al. (2026) show that LLM representations broadly recover human category boundaries, while failing to reflect fine-grained typicality structur

← Previous
1…439440441442443…1060
Next →