Archives

June 2026

2731 articles

GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> addyosmani /</span> agent-skills

Agent-skills: open-source framework to equip AI coding agents with production-grade engineering capabilities. GitHub repository aiming to standardize coding agent competencies.

AI AgentsCode generationOpen source
SIG
45
HYP
35
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> x1xhlol /</span> system-prompts-and-models-of-ai-tools

GitHub repository collecting system prompts and internal models from 25+ AI tools (Claude Code, Cursor, Devin AI, Perplexity, Replit, v0, etc.). Includes open-source alternatives. Resource to reverse-engineer system instructions of popular assistants.

Claude CodeCode generationPrompt engineering
SIG
45
HYP
55
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> Ataraxy-Labs /</span> sem

Ataraxy-Labs/sem: semantic version control on top of git with entity-level diffs, blame, and impact analysis. Supports 26 languages via tree-sitter. Built for coding agents.

AI AgentsCode generationTools
SIG
72
HYP
25
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> cube-js /</span> cube

Cube Core is an open-source semantic layer for AI, BI and embedded analytics. The project is gaining traction on GitHub Trending.

Open sourceInfrastructureRAG
SIG
45
HYP
35
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> chroma-core /</span> chroma

Chroma is a vector search infrastructure for AI applications. The trending GitHub project provides storage and querying tools for embeddings to support RAG and language model-based systems.

Vector searchEmbeddingsRAG
SIG
65
HYP
25
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> luongnv89 /</span> asm

Asm is a universal skill manager for AI coding agents. The GitHub project provides infrastructure to orchestrate and manage capabilities of autonomous coding agents.

AI AgentsCode generationTools
SIG
35
HYP
45
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> wonderwhy-er /</span> DesktopCommanderMCP

DesktopCommanderMCP is an MCP server for Claude providing terminal control, file system search, and diff-based file editing capabilities.

ClaudeMCPTools
SIG
65
HYP
20
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> wanshuiyin /</span> Auto-claude-code-research-in-sleep

ARIS (Auto-Research-In-Sleep): lightweight Markdown-based framework for autonomous ML research. Cross-model review loops, idea discovery, and experiment automation. Works with Claude Code, Codex, OpenClaw, and any LLM agent.

Claude CodeAI AgentsMulti-agent
SIG
45
HYP
55
GitHub Trending·

<svg aria-hidden="true" data-component="Octicon" height="16" viewBox="0 0 16 16" version="1.1" width="16" data-view-component="true" class="octicon octicon-repo mr-1 tmp-mr-1 color-fg-muted"> <path d="M2 2.5A2.5 2.5 0 0 1 4.5 0h8.75a.75.75 0 0 1 .75.75v12.5a.75.75 0 0 1-.75.75h-2.5a.75.75 0 0 1 0-1.5h1.75v-2h-8a1 1 0 0 0-.714 1.7.75.75 0 1 1-1.072 1.05A2.495 2.495 0 0 1 2 11.5Zm10.5-1h-8a1 1 0 0 0-1 1v6.708A2.486 2.486 0 0 1 4.5 9h8ZM5 12.25a.25.25 0 0 1 .25-.25h3.5a.25.25 0 0 1 .25.25v3.25a.25.25 0 0 1-.4.2l-1.45-1.087a.249.249 0 0 0-.3 0L5.4 15.7a.25.25 0 0 1-.4-.2Z"></path> </svg> <span data-view-component="true" class="text-normal"> anthropics /</span> claude-code-security-review

GitHub Action powered by Claude that automatically analyzes code changes to detect security vulnerabilities.

ClaudeCode generationTools
SIG
65
HYP
25
arXiv cs.AI·

Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy

AI-MASLD, a stress-audit framework, evaluates 7 medical LLMs on 240 clinical cases with narrative perturbations. All perform well at baseline but diverge under realistic stress. Quantized models hide functional collapse; medical fine-tuning degrades logical stability and fairness. An open-weight model matches or exceeds proprietary alternatives on all safety dimensions.

BenchmarksAI safetyEvals
SIG
78
HYP
15
arXiv cs.AI·

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

MemToolAgent improves LLM agent tool use through structured memory management. The framework extracts past experiences into memory entries, dynamically retrieves relevant ones, and generates reflections from user feedback. Achieves 29%, 80%, and 17% relative improvements on WorkBench, NESTFUL, and PEToolBench without fine-tuning.

AI AgentsToolsBenchmarks
SIG
72
HYP
28
arXiv cs.LG·

DiffoR: A Unified Continuous Generative Framework for Universal Ordinal Regression

DiffOR introduces a novel paradigm for Ordinal Regression as Continuous Generative task. The framework leverages diffusion models to recover continuous ordinal values via iterative denoising, with a Dual-Decoupling Strategy (Multi-scale Increment Aggregation and Dynamic Denoising Perception) to preserve ordinal topology. Validated on 12 benchmarks across four domains.

PapersBenchmarksReasoning
SIG
72
HYP
28
arXiv cs.LG·

MST-Direct at Scale: Multivariate and Conditional Geostatistical Simulation via Sinkhorn Optimal Transport

MST-Direct scaled to large-scale multivariate and conditional geostatistical simulation via Sinkhorn optimal transport. Addresses scalability (O(nC) memory), multivariate extension, and kriging-based conditioning. Validated on 6-variate heteroscedastic distribution, grids 200×200 and 100×100 with 200 hard-data samples. Reproduces joint distribution exactly versus PPMT approximation.

PapersBenchmarks
SIG
72
HYP
08
arXiv cs.LG·

STARIXNet: Multivariate and Multi-attribute Deep Learning Approach to Real-Time Resource Allocation in Cloud Platforms

STARIXNet is a lightweight neural network for real-time resource allocation in cloud platforms. It captures spatio-temporal relationships among multiple system metrics (seasonality, trend, auto-regression, exogenous variables) and prioritizes service stability over forecast accuracy. Deployed at Walmart, it achieves 10-50% cost savings.

InfrastructureBenchmarksCode generation
SIG
75
HYP
25
arXiv cs.LG·

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

RL4F is an open-source offline reinforcement learning benchmark for plasma control in nuclear fusion. Built on historical data from the DIII-D tokamak, it evaluates imitation learning and offline RL methods on four multi-actuator tracking tasks (rotation, density, temperature, pressure). Offline model-based RL methods achieve best average performance.

Reinforcement learningBenchmarksOpen source
SIG
82
HYP
15
arXiv cs.LG·

Measuring Poverty and Inequality with Reduced Data: A Machine Learning Approach Using Nigerian Household Data

Study applying Random Forest Recursive Feature Elimination to Nigerian household survey data (2018/19) to identify minimal predictors of poverty status, welfare quintiles, and inequality. RF-RFE achieves 90% accuracy for poverty with 5 income variables, 80% for seasonal quintiles. ML methods reduce data requirements while preserving distributional information for poverty and inequality monitoring.

BenchmarksPapers
SIG
72
HYP
15
arXiv cs.LG·

QDSP: An Interpretable Structured Learning Framework for Predicting Death or Cerebral Palsy in Very Low Birth Weight Infants

QDSP is a structured learning framework to predict mortality or cerebral palsy in very low birth weight infants. On a cohort of 51 patients, it achieves 92% accuracy and 0.9714 AUC, outperforming XGBoost, TabNet, and TabPFN. Interpretability via SHAP identifies clinically relevant predictors including cystic periventricular leukomalacia.

BenchmarksEvalsAI safety
SIG
72
HYP
15
arXiv cs.LG·

Reachability and asymptotics of Gaussian Transformer dynamics

Theoretical study formalizes data propagation through Transformers as a nonlinear control system. For mean-field Transformer with self-attention and affine feed-forward layers, Gaussian distributions remain exactly Gaussian. This reduces dynamics to finite-dimensional bilinear control system governing mean and covariance evolution, connecting Transformer expressivity to Riccati-type equations.

PapersReasoningBenchmarks
SIG
78
HYP
15
arXiv cs.LG·

UNIQ: Conformal Calibration for Adaptive Conservatism in Offline Reinforcement Learning

UNIQ introduces conformal calibration for adaptive conservatism in offline reinforcement learning. Built on IQL, the method uses a multi-expectile ensemble and split conformal prediction for distribution-free uncertainty estimation, dynamically adjusting penalties based on local data coverage. On D4RL MuJoCo, UNIQ outperforms IQL with 10× lower memory than EDAC.

Reinforcement learningPapersBenchmarks
SIG
78
HYP
15