Nightly digest · AI + CS signal

23 JUL 2026

archived edition221 fetched → 29 published

Tonight’s top picks

read these, skip the rest

ResearcharXiv cs.CLarxiv.org2026-07-22

PyroDash is a cost-aware framework for token-level SLM-LLM collaborative inference.

  • SLM
  • LLM
  • token-level

PyroDash trains the SLM in three stages and balances answer accuracy against inference cost. It requires neither a separate router nor LLM retraining. The framework is useful for applications where both speed and accuracy are crucial.

ResearcharXiv cs.CRarxiv.org2026-07-22

Orchestrated ensembles of small language models can match single LLM performance on malware analysis.

  • 11 SLMs
  • 3 pre-trained
  • 6 LLMs

The study investigates the effectiveness of small language models in malware analysis, which is essential for cybersecurity. The results show that ensembles of small models can be a viable alternative to large language models.

ResearcharXiv cs.CLarxiv.org2026-07-22

LKValues is a resource suite for Sri Lankan value alignment in large language models.

  • Sri Lankan
  • 40 values
  • 150k instances

LKValues addresses the cultural bias in large language models by providing a survey-grounded resource suite for Sri Lankan value alignment. This is crucial for developing culturally sensitive AI models.

ContestCodeforcescodeforces.com2026-08-01

Codeforces Round starts August 1

  • August 1
  • 2.0h

The contest will start on August 1 and will last for 2 hours.

ResearcharXiv quant-pharxiv.org2026-07-22

Qoreo programming language introduced

Qoreo is a choreographic programming language for quantum distributed systems. It includes a local quantum language with linear types and a choreographic language that combines local quantum computation with inter-actor classical and quantum communication.

10 items

ResearcharXiv cs.CLarxiv.org2026-07-22

PyroDash is a cost-aware framework for token-level SLM-LLM collaborative inference.

  • SLM
  • LLM
  • token-level

PyroDash trains the SLM in three stages and balances answer accuracy against inference cost. It requires neither a separate router nor LLM retraining. The framework is useful for applications where both speed and accuracy are crucial.

ResearcharXiv cs.CRarxiv.org2026-07-22

Orchestrated ensembles of small language models can match single LLM performance on malware analysis.

  • 11 SLMs
  • 3 pre-trained
  • 6 LLMs

The study investigates the effectiveness of small language models in malware analysis, which is essential for cybersecurity. The results show that ensembles of small models can be a viable alternative to large language models.

ResearcharXiv cs.CLarxiv.org2026-07-22

LKValues is a resource suite for Sri Lankan value alignment in large language models.

  • Sri Lankan
  • 40 values
  • 150k instances

LKValues addresses the cultural bias in large language models by providing a survey-grounded resource suite for Sri Lankan value alignment. This is crucial for developing culturally sensitive AI models.

NewsHacker Newsgithub.com427 pts2026-07-22

GigaToken is a language model tokenization system.

  • ~1000x faster

GigaToken aims to improve the efficiency of language model tokenization, which is a critical component of natural language processing. Faster tokenization can lead to better performance and reduced latency.

Newsopenai blogopenai.com2026-07-22

OpenAI Presence is an enterprise AI agent platform.

  • voice
  • chat
  • agents

OpenAI Presence is designed to help organizations deploy trusted AI agents for customer and internal workflows. The platform aims to provide a reliable and efficient way to integrate AI into business operations.

Newsdeepmind blogdeepmind.google2026-07-22

Google commits $40M to the Genesis Mission.

  • $40M
  • Genesis Mission

The commitment is part of Google's effort to accelerate scientific discovery using AI. The Genesis Mission aims to advance our understanding of the world through cutting-edge research and innovation.

Newsopenai blogopenai.com2026-07-22

OpenAI outlines its commitment to advancing American science.

  • U.S. Department of Energy
  • national labs

OpenAI's commitment is focused on using frontier AI to accelerate discovery and drive innovation in the scientific community. The partnership with the U.S. Department of Energy and national labs aims to leverage AI for the betterment of society.

Newsopenai blogopenai.com2026-07-22

OpenAI announces Project Camellia in Effingham County, Georgia.

  • Effingham County
  • Georgia
  • Project Camellia

Project Camellia is part of OpenAI's effort to promote responsible energy, community investment, and job creation. The project aims to provide access to Codex and other AI resources to the local community.

NewsAnthropicanthropic.com

Anthropic donates $20 million to Public First Action.

  • $20 million
  • Public First Action

The donation is part of Anthropic's commitment to supporting organizations that align with its values and mission. Public First Action is a non-profit organization focused on promoting public interest and social welfare.

AnalysisAnthropicanthropic.com

Anthropic outlines a research agenda for the Economic Futures Research Fund.

  • Economic Futures
  • research agenda

The research agenda is focused on exploring the potential applications and implications of AI in the economic sector. The goal is to provide insights and recommendations for policymakers and stakeholders.

6 items

Newslwnlwn.net2026-07-22

Security updates have been issued by multiple Linux distributions.

  • AlmaLinux
  • Debian
  • Fedora

The security updates address various vulnerabilities and issues in different Linux distributions. Users are advised to apply the updates to ensure the security and stability of their systems.

Newslwnlwn.net2026-07-22

BPF programs can now be attached to multiple tracepoints.

  • BPF
  • tracepoints
  • kernel 7.2

The update allows for more flexible and efficient use of BPF programs in kernel tracing and debugging. This is a significant improvement for developers and system administrators.

Newslwnlwn.net2026-07-23

LWN.net Weekly Edition for July 23, 2026 released

  • July 23, 2026

The edition covers various topics including LLMs in the kernel, GNOME save and restore, and BPF and tracepoints. It also includes briefs on GNOME security, PyPI policy, and Firefox 153.

Newslwnlwn.net2026-07-22

PyPI rejects new files after 14 days

  • 14 days

The restriction is to prevent the poisoning of old releases if publishing tokens or workflows of PyPI projects are compromised. This change was made after the popular packages LiteLLM and Telnyx were compromised.

Newslwnlwn.net2026-07-22

Save and restore may come to GNOME

  • GNOME 51
  • October

The feature is expected to provide a platform-wide save and restore framework for GNOME. It has been attempted twice before but is expected to succeed this time.

hackadayhackaday.com2026-07-22

LLM in the kitchen

The article discusses the use of LLMs in recipe search and how it can help with finding specific details in a recipe.

1 items

ContestCodeforcescodeforces.com2026-08-01

Codeforces Round starts August 1

  • August 1
  • 2.0h

The contest will start on August 1 and will last for 2 hours.

6 items

ResearcharXiv quant-pharxiv.org2026-07-22

Qoreo programming language introduced

Qoreo is a choreographic programming language for quantum distributed systems. It includes a local quantum language with linear types and a choreographic language that combines local quantum computation with inter-actor classical and quantum communication.

ResearcharXiv cs.DBarxiv.org2026-07-22

Hollywood movie dataset introduced

  • 200,000
  • 19.7M

Hollywood is a synthetic IMDb-compatible benchmark generator that combines LLM-generated semantic dictionaries with deterministic temporal-graph-based relational data generation. It contains 200,000 primary movies and 19.7M IMDb-style rows.

ResearcharXiv cs.LOarxiv.org2026-07-22

Linearising explicit substitutions using intersection types

The paper defines a new term expansion for a calculus with explicit substitutions, using it to relate a lambda-calculus with explicit substitutions to Boudol's resource aware lambda-calculus with multiplicities.

ResearcharXiv cs.DCarxiv.org2026-07-22

Odin synchronization system introduced

Odin is a distributed point-based neural rendering training system that replaces global barriers with primitive-level synchronization. It uses a ahead-of-time scheduler and runtime validation to identify low-conflict overlap windows.

ResearcharXiv cs.DBarxiv.org2026-07-22

RCC speculative write versioning introduced

RCC is a system that leverages redo logs to resolve conflicting writes. It allows concurrent transactions to be pipelined after the update-conflicting point, enabling lightweight rollback.

ResearcharXiv cs.CLarxiv.org2026-07-22

TriAgent multi-agent committee introduced

  • 0.87 F1

TriAgent is a multi-agent committee stratified by contextual granularity for cost-efficient financial sentiment analysis. It uses a three-way Semantic Divergence Index to route queries accordingly.

6 items

RepoGitHub trending · pythongithub.com2,100 pts

QuantMind framework introduced

QuantMind is an agent-native knowledge extraction and retrieval framework for quantitative finance.

RepoGitHub trending · pythongithub.com3,294 pts

NVIDIA's Model-Optimizer library optimizes deep learning models

  • SOTA techniques
  • TensorRT-LLM

It supports various downstream deployment frameworks and compresses models for faster inference speed. The library includes techniques like quantization and neural architecture search.

RepoGitHub trending · rustgithub.com4,363 pts

Open-source AI-native vector design tool OpenPencil released

It features concurrent Agent Teams and allows designing UI directly on the live canvas. OpenPencil is a modern alternative to Pencil.

RepoGitHub trending · c++github.com13,080 pts

Assimp library loads 40+ 3D-file-formats

  • 40+ formats

The library provides a unified and clean data structure for loaded 3D models. It is useful for various applications that require 3D model loading.

RepoGitHub trending · c++github.com21,286 pts

Catch2 is a modern C++ test framework

  • C++14
  • C++17

It supports unit-tests, TDD, and BDD. Catch2 also has support for older C++ versions in separate branches.

RepoGitHub trending · c++github.com29,241 pts

Spdlog is a fast C++ logging library

It is designed to be highly performant and efficient. Spdlog can be used in various applications that require fast logging capabilities.

fetch → dedupe → score → summarize, nightly via GitHub Actionsj/k navigate · enter expands · o opens sourcepast editions →