Tech News
I found 10k GitHub repositories distributing Trojan malware
A security researcher discovered over 10,000 GitHub repositories distributing Trojan malware, often by cloning legitimate projects and adding malicious code. The repos are updated frequently to appear in search result.
Datasette Apps: Host custom HTML applications inside Datasette
The update describes today we launched a new plugin for Datasette, datasette-apps , with this launch announcement post on the Datasette project blog. That post has the what , but I'm going to expand.
Amazon Bedrock AgentCore harness is now generally available: Go from idea to production-grade agent in minutes
This item points to a practical AI tooling or research update worth checking from the source.
Using AI to help physicians diagnose rare genetic diseases affecting children
The update describes researchers used an OpenAI reasoning model to help diagnose rare diseases, identifying 18 new diagnoses in previously unsolved cases.
Building AI Agents for AR Glasses and XR Devices with NVIDIA XR AI
The update describes developers building for AR glasses and wearable devices face an infrastructure gap. The hardware is ready, but creating AI experiences requires integrating live.
GitHub Repos
sansan0/TrendRadar
ClassicPythonaibarkunslothai/unsloth
ClassicPythonagentdeepseektaylorwilsdon/google_workspace_mcp
High PotentialPythonaig-suiteResearch Papers
RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
RM-Bench is a new benchmark that tests reward models on their ability to detect subtle content differences and resist style biases, revealing that even top models perform near random when style is varied.
Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows
Spider 2.0 is a new benchmark for evaluating language models on real-world enterprise text-to-SQL tasks, featuring complex databases, multiple SQL dialects, and long workflows. Current models solve only 21.3% of tasks, highlighting a large gap from practical deployment.
StylisticBias: A Few Human Visual Cues Drive Most Social Biases in MLLMs
Multimodal large language models (MLLMs) are increasingly deployed in personally and societally consequential settings, yet the visual cues that shape how these models judge people remain poorly understood. Prior work often compares different (groups of) individuals, making.
