AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,563 results
7 May 2026

A Comparative Study of PyCaret AutoML and CNN-BiLSTM for Binary Hate Speech Detection in Indonesian Twitter

Model ReleasesDGX agent

arXiv:2605.04885v1 Announce Type: new Abstract: This paper compares a PyCaret AutoML branch and a CNN-BiLSTM branch for binary hate speech detection on Indonesian Twitter using the HS label from the c

A few weeks ago @simonw got Claude to port LiteParse to the browser. Today, we are launching that work as a complete guide in our docs! http…

Model ReleasesDGX agent

A few weeks ago @simonw got Claude to port LiteParse to the browser. Today, we are launching that work as a complete guide in our docs! https://developers.llamaindex.ai/liteparse/guides/browser-usage/

A Regulatory Governance Framework for AI-Driven Financial Fraud Detection in U.S. Banking: Integrating OCC, SR 11-7, CFPB, and FinCEN Compliance Requirements for Model Development, Validation, and Monitoring Lifecycles


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2605.04076v1 Announce Type: new Abstract: U.S. financial institutions deploying AI-based fraud detection face a fragmented compliance landscape spanning four regulatory frameworks -- OCC Bulleti

A Scalable Multi-Task Model for Virtual Sensors

Model ReleasesDGX agent

arXiv:2601.20634v2 Announce Type: replace Abstract: Virtual sensors replace expensive physical sensors in critical applications through machine learning by predicting target signals from available mea

A Self-Attentive Meta-Optimizer with Group-Adaptive Learning Rates and Weight Decay

Model ReleasesDGX agent

arXiv:2605.04055v1 Announce Type: new Abstract: Adaptive optimizers like AdamW apply uniform hyperparameters across all parameter groups, ignoring heterogeneous optimization dynamics across layers and

A unified Benchmark for Multi-Frame Image Restoration under Severe Refractive Warping

Model ReleasesDGX agent

arXiv:2605.05079v1 Announce Type: new Abstract: Video sequence capturing through refractive dynamic media, such as a turbulent air or water surface, often suffer from severe geometric distortions and

A Universal Large Language Model -- Drone Command and Control Interface

Model ReleasesDGX agent

arXiv:2601.15486v2 Announce Type: replace Abstract: The use of artificial intelligence (AI) for drone control can have a transformative impact on drone capabilities, especially when real world informa

Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkir

Model ReleasesDGX agent

arXiv:2605.04948v1 Announce Type: new Abstract: This paper presents a comparative study of parameter-efficient fine-tuning (PEFT) methods, including LoRA and QLoRA, applied to the task of adapting lar

Adaptive Ensemble Aggregation for Actor-Critics

Model ReleasesDGX agent

arXiv:2507.23501v2 Announce Type: replace Abstract: Ensembles are ubiquitous in off-policy actor-critic learning, yet their efficacy depends critically on how they are aggregated. Current methods typi

Advancing voice intelligence with new models in the API

Model ReleasesDGX agent

OpenAI announced new voice intelligence models available through its API, expanding capabilities for developers to integrate advanced voice processing and understanding features into their application

Aes3D: Aesthetic Assessment in 3D Gaussian Splatting

Model ReleasesDGX agent

arXiv:2605.05155v1 Announce Type: new Abstract: As 3D Gaussian Splatting (3DGS) gains attention in immersive media and digital content creation, assessing the aesthetics of 3D scenes becomes important

Agent harnesses have an expiration date

Model ReleasesDGX agent

A benchmark-driven look at why agent harnesses need adaptive finish logic as model behavior changes across Claude, GPT-4o, and Gemma. The post Agent harnesses have an expiration date appeared first on

Agentic Vulnerability Reasoning on Windows COM Binaries

Model ReleasesDGX agent

arXiv:2605.05000v1 Announce Type: cross Abstract: Windows Component Object Model (COM) services run with elevated privileges and are widely accessible to authenticated users, making race conditions in

Aggressive or Imperceptible, or Both: Network Pruning Assisted Hybrid Byzantines in Federated Learning

Model ReleasesDGX agent

arXiv:2404.06230v3 Announce Type: replace Abstract: In federated learning (FL), profiling and verifying each client is inherently difficult, which introduces a significant security vulnerability: mali

Anthropic is letting Claude agents ‘dream’ so they don’t sleep on the job

Model ReleasesDGX agent

Anthropic PBC said today it’s giving its AI agents the ability to “dream” and remember past interactions and work they’ve performed so they can identify recurring mistakes and improve over time. In an

Anthropic just shipped sleep into agents. When you sleep, your hippocampus replays the day's neural sequences to the cortex during 150-220 H…

Model ReleasesDGX agent

Anthropic just shipped sleep into agents. When you sleep, your hippocampus replays the day's neural sequences to the cortex during 150-220 Hz bursts called sharp-wave ripples. The replay runs about 20

Anthropic researchers detail 'natural language autoencoders', which convert LLM activations, the numbers encoding a model's thoughts, into natural language text (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic researchers detail “natural language autoencoders”, which convert LLM activations, the numbers encoding a model's thoughts, into natural language text — When you talk to an AI mod

Are LLMs Ready for Conflict Monitoring? Empirical Evidence from West Africa

Model ReleasesDGX agent

arXiv:2605.04177v1 Announce Type: new Abstract: As LLMs enter conflict monitoring, understanding systematic distortions in their outputs is critical for humanitarian accountability. We evaluate four v

Are Multimodal LLMs Ready for Clinical Dermatology? A Real-World Evaluation in Dermatology

Model ReleasesDGX agent

arXiv:2605.04098v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated promise on publicly available dermatology benchmarks. However, benchmark performance may not

ARTA: Adversarial-Robust Multivariate Time--Series Anomaly Detection via Sparsity-Constrained Perturbations

Model ReleasesDGX agent

arXiv:2603.25956v2 Announce Type: replace Abstract: Time-series anomaly detection (TSAD) is a critical component in monitoring complex systems, yet modern deep learning-based detectors are often highl

Assessing Cognitive Effort in L2 Idiomatic Processing: An Eye-Tracking Dataset

Model ReleasesDGX agent

arXiv:2605.04857v1 Announce Type: new Abstract: This paper presents the development and validation of an eye-tracking dataset designed to investigate how second-language (L2) learners process idiomati

AsymmetryZero: A Framework for Operationalizing Human Expert Preferences as Semantic Evals

Model ReleasesDGX agent

arXiv:2605.04083v1 Announce Type: new Abstract: Much of the focus in RL today is on evaluation design: building meaningful evals that serve simultaneously as benchmarks and as well-defined reward sign

Automated Formal Proofs of Combinatorial Identities via Wilf-Zeilberger Guidance and LLMs

Model ReleasesDGX agent

arXiv:2605.04472v1 Announce Type: new Abstract: Automating formal proofs of combinatorial identities is challenging for LLM-based provers, as long-horizon proof planning is required and unconstrained

Automated Large-scale CVRP Solver Design via LLM-assisted Flexible MCTS

Model ReleasesDGX agent

arXiv:2605.03339v1 Announce Type: new Abstract: Solving large-scale CVRP (LSCVRP) with hundreds to thousands of nodes remains difficult for even state-of-the-art solvers. Divide-and-conquer can scale

Back to Blackwell: Closing the Loop on Intransitivity in Multi-Objective Preference Fine-Tuning

Model ReleasesDGX agent

arXiv:2602.19041v2 Announce Type: replace Abstract: A recurring challenge in preference fine-tuning (PFT) is handling extit{intransitive} (i.e., cyclic) preferences. Intransitive preferences often ste

Bayesian Parameter Shift Rule in Variational Quantum Eigensolvers

Model ReleasesDGX agent

arXiv:2502.02625v2 Announce Type: replace Abstract: Parameter shift rules (PSRs) are key techniques for efficient gradient estimation in variational quantum eigensolvers (VQEs). In this paper, we prop

Behind the Scenes Hardening Firefox with Claude Mythos Preview

Model ReleasesDGX agent

Behind the Scenes Hardening Firefox with Claude Mythos Preview Fascinating, in-depth details on how Mozilla used their access to the Claude Mythos preview to locate and then fix hundreds of vulnerabil

Benchmarking LLMs on the Massive Sound Embedding Benchmark (MSEB)

Model ReleasesDGX agent

arXiv:2605.04556v1 Announce Type: cross Abstract: The Massive Sound Embedding Benchmark (MSEB) has emerged as a standard for evaluating the functional breadth of audio models. While initial baselines

Benchmarking POS Tagging for the Tajik Language: A Comparative Study of Neural Architectures on the TajPersParallel Corpus

Model ReleasesDGX agent

arXiv:2605.04576v1 Announce Type: new Abstract: This paper presents the first benchmark for the task of automatic part-of-speech (POS) tagging for the Tajik language. Despite the existence of multilin

BenCSSmark: Making the Social Sciences Count in LLM Research

Model ReleasesDGX agent

arXiv:2605.04886v1 Announce Type: new Abstract: This position paper argues that the under-representation of social science tasks in contemporary LLM benchmarks limits advances in both LLM evaluation a

Budgeted LoRA: Distillation as Structured Compute Allocation for Efficient Inference

Model ReleasesDGX agent

arXiv:2605.04341v1 Announce Type: cross Abstract: We study distillation for large language models under explicit compute constraints, with the goal of producing student models that are not only cheape

Builders at Code with Claude

Model ReleasesDGX agent

'Builders at Code with Claude' likely covers best practices and techniques for developers using Claude to build and code applications, possibly including practical examples, tips for effective prompti

Capacity-Aware Mixture Law Enables Efficient LLM Data Optimization

Model ReleasesDGX agent

arXiv:2603.08022v2 Announce Type: replace Abstract: A data mixture refers to how different data sources are combined to train large language models, and selecting an effective mixture is crucial for o

CARD: A Multi-Modal Automotive Dataset for Dense 3D Reconstruction in Challenging Road Topography

Model ReleasesDGX agent

arXiv:2605.05014v1 Announce Type: new Abstract: Autonomous driving must operate across diverse surfaces to enable safe mobility. However, most driving datasets are captured on well-paved flat roads. M

Claude Code just stopped a DDoS attack on BridgeMind in under 10 minutes. 13 million requests per minute hitting our API. CPU pegged at 94%.…

Model ReleasesDGX agent

Claude Code just stopped a DDoS attack on BridgeMind in under 10 minutes. 13 million requests per minute hitting our API. CPU pegged at 94%. Latency spiking to 60 seconds. Production was down. I opene

Claude for Excel, PowerPoint, and Word are now generally available, and Claude for Outlook is in public beta. As Claude moves between your M…

Model ReleasesDGX agent

Claude for Excel, PowerPoint, and Word are now generally available, and Claude for Outlook is in public beta. As Claude moves between your Microsoft apps, it carries the full context of your conversat

Closed-Loop Vision-Language Planning for Multi-Agent Coordination

Model ReleasesDGX agent

arXiv:2502.10148v3 Announce Type: replace Abstract: Cooperative multi-agent reinforcement learning (MARL) struggles with sample efficiency, interpretability, and generalization. While Large Language M

Coding plan users interested in early experimentation can fill out this form: http://docs.google.com/forms/d/e/1FAIpQLSdEg9C_7FRQWRbnJt--BJX…

Model ReleasesDGX agent

Zhipu AI is inviting users interested in early experimentation with its coding capabilities to register through a Google Form. This form likely allows developers or researchers to gain access to beta

Cognitive Twins: Investigating Personalized Thinking Model Building and Its Performance Enhancement with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.04761v1 Announce Type: new Abstract: This paper presents the Personalized Thinking Model (PTM), a hierarchical and interpretable learner representation designed for AI supported education.

Computer-Aided Design Generation by Cascaded Discrete Diffusion Model

Model ReleasesDGX agent

arXiv:2605.05031v1 Announce Type: new Abstract: Recent deep learning approaches seek to automate CAD creation by representing a model as a sequence of discrete commands and parameters, and then genera

Conceptors for Semantic Steering

Model ReleasesDGX agent

arXiv:2605.04980v1 Announce Type: cross Abstract: Activation-based steering provides control of LLM behavior at inference time, but the dominant paradigm reduces each concept to a single direction who

Conflict-Aware Fusion: Mitigating Logic Inertia in Large Language Models via Structured Cognitive Priors

Model ReleasesDGX agent

arXiv:2512.06393v5 Announce Type: replace-cross Abstract: Large language models (LLMs) achieve high accuracy on many reasoning benchmarks but remain brittle under structural perturbations of rule-base

ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2605.05126v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models primarily focus on mapping 2D observations to actions, but exhibit notable limitations in spatiotemporal per

Constrained Extreme Gradient Boosting for Adapting Reduced-Order Models

Model ReleasesDGX agent

arXiv:2605.04130v1 Announce Type: new Abstract: High-fidelity simulations, such as computational fluid dynamics and finite element analysis, are essential for modeling complex engineering systems but

Constraint-Enhanced Reinforcement Learning Based on Dynamic Decoupled Spherical Radial Squashing

Model ReleasesDGX agent

arXiv:2605.04185v1 Announce Type: new Abstract: When deploying reinforcement learning policies to physical robots, actuator rate constraints -- hard limits on how fast each joint can move per control

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live

Model ReleasesDGX agent

arXiv:2511.02230v4 Announce Type: replace-cross Abstract: KV cache management is essential for efficient LLM inference. To maximize utilization, existing inference engines evict finished requests' KV

Cross-Model Consistency of Feature Importance in Electrospinning: Separating Robust from Model-Dependent Features

Model ReleasesDGX agent

arXiv:2605.04905v1 Announce Type: new Abstract: Electrospinning is a highly sensitive fabrication process in which small variations in operating parameters can significantly influence fiber morphology

DALight-3D: A Lightweight 3D U-Net for Brain Tumor Segmentation from Multi-Modal MRI

Model ReleasesDGX agent

arXiv:2605.04518v1 Announce Type: new Abstract: Automatic brain tumor segmentation from multi-modal MRI remains challenging because volumetric models often incur substantial computational cost. This p

DART: A Vision-Language Foundation Model for Comprehensive Rope Condition Monitoring

Model ReleasesDGX agent

arXiv:2605.04943v1 Announce Type: new Abstract: The condition monitoring (CM) of synthetic fibre ropes (SFRs) used in offshore, maritime, and industrial settings demands more than a classifier: inspec

Deep Reprogramming Distillation for Medical Foundation Models

Model ReleasesDGX agent

arXiv:2605.04447v1 Announce Type: new Abstract: Medical foundation models pre-trained on large-scale datasets have shown powerful versatile performance. However, when adapting medical foundation model

Delta-Based Neural Architecture Search: LLM Fine-Tuning via Code Diffs

Model ReleasesDGX agent

arXiv:2605.04903v1 Announce Type: cross Abstract: Large language models (LLMs) show strong potential for neural architecture generation, yet existing approaches produce complete model implementations

Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone

Model ReleasesDGX agent

arXiv:2605.04454v1 Announce Type: cross Abstract: Alignment evaluation in machine learning has largely become evaluation of models. Influential benchmarks score model outputs under fixed inputs, such

DiffCap-Bench: A Comprehensive, Challenging, Robust Benchmark for Image Difference Captioning

Model ReleasesDGX agent

arXiv:2605.04503v1 Announce Type: new Abstract: Image Difference Captioning (IDC) generates natural language descriptions that precisely identify differences between two images, serving as a key bench

Differentiable Chemistry in PINNs for Solving Parameterized and Stiff Reaction Systems

Model ReleasesDGX agent

arXiv:2605.04708v1 Announce Type: new Abstract: From neural ODEs to continuous-time machine learning, differentiable solvers allow physics, optimization, and simulation to become trainable components

Discovering New Theorems via LLMs with In-Context Proof Learning in Lean

Model ReleasesDGX agent

arXiv:2509.14274v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated significant promise in formal theorem proving. In this study, we investigate the ability of LLMs to d

Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout

Model ReleasesDGX agent

arXiv:2605.05092v1 Announce Type: cross Abstract: Safe L2/L3 driving automation requires anticipating human-in-the-loop reactions during shared-control transitions. While most driving world models for

Elon Musk says SpaceX reserves 'the right to reclaim the compute' from Anthropic if its 'AI engages in actions that harm humanity' (Elon Musk/@elonmusk)

Model ReleasesDGX agent

Elon Musk / @elonmusk: Elon Musk says SpaceX reserves “the right to reclaim the compute” from Anthropic if its “AI engages in actions that harm humanity” — @MobofJoggers @nottombrown Just as SpaceX la

Emergent Hierarchical Structure in Large Language Models: An Information-Theoretic Framework for Multi-Scale Representation

Model ReleasesDGX agent

arXiv:2505.18244v3 Announce Type: replace Abstract: Why do language models from different architecture families respond so differently to the same perturbation? We argue that the answer is not scale,

Empirical Study of Pop and Jazz Mix Ratios for Genre-Adaptive Chord Generation

Model ReleasesDGX agent

arXiv:2605.04998v1 Announce Type: cross Abstract: Chord progression generation is practically important but understudied. Most large-scale symbolic music systems target melody, multi-track arrangement

Enhancing Agent Safety Judgment: Controlled Benchmark Rewriting and Analogical Reasoning for Deceptive Out-of-Distribution Scenarios

Model ReleasesDGX agent

arXiv:2605.03242v1 Announce Type: new Abstract: Tool-using agent systems powered by large language models (LLMs) are increasingly deployed across web, app, operating-system, and transactional environm

← Previous
1…277278279280281…377
Next →