{"articles":[{"slug":"deepseek-elastic-compute-dsec","title":"DeepSeek Elastic Compute (DSec)","subtitle":null,"summary":"# Computer Science > Distributed, Parallel, and Cluster Computing [Submitted on 19 Sep 2026] # Title:DeepSeek Elastic Compute (DSec): A Sandbox Infrastructure for Effective Agentic Training at Scale View PDF HTML (experimental) Abstract:Large-scale agentic training and evaluation with large language models (LLMs) rely on isolated, stateful execution environments in which models inspect repositories, invoke tools, execute commands, and interact with task-specific services. These workloads create sandboxes in large bursts, span heterogeneous functionality and isolation…","content_type":"research","language":"en","canonical_url":"https://arxiv.org/abs/2609.22978","author":{"name":"DeepSeek-AI","url":null,"person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"DeepSeek","url":"https://www.deepseek.com/","listing_slug":"deepseek","listing":{"slug":"deepseek","name":"DeepSeek","listing_type":"company","url":"https://listedstartups.com/companies/deepseek"}},"topics":[{"name":"AI","slug":"ai","url":"https://listedarticles.com/topics/ai"},{"name":"Machine Learning","slug":"machine-learning","url":"https://listedarticles.com/topics/machine-learning"},{"name":"Research","slug":"research","url":"https://listedarticles.com/topics/research"},{"name":"Infrastructure","slug":"infrastructure","url":"https://listedarticles.com/topics/infrastructure"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":3693,"reading_minutes":16,"published_at":"2026-09-26T12:00:00.000Z","added_at":"2026-09-27T00:17:53.167Z","updated_at":"2026-09-27T00:17:53.167Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":true},"profile_url":"https://listedarticles.com/articles/deepseek-elastic-compute-dsec","markdown_url":"https://listedarticles.com/articles/deepseek-elastic-compute-dsec.md","example":false,"citation":"DeepSeek-AI, DeepSeek. \"DeepSeek Elastic Compute (DSec).\" 26 Sept 2026. https://arxiv.org/abs/2609.22978 (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://arxiv.org/abs/2609.22978"},"snippet":null,"score":null},{"slug":"mercury-2-5-intelligence-performance-and-price-analysis","title":"Mercury 2.5: Intelligence, Performance and Price Analysis","subtitle":null,"summary":"Artificial Analysis profiles Inception's Mercury 2.5—Intelligence Index, ~770 output tokens/sec, pricing, and where the diffusion LLM sits on the quality-vs-speed frontier.","content_type":"research","language":"en","canonical_url":"https://artificialanalysis.ai/models/mercury-2-5","author":{"name":"Artificial Analysis","url":null,"person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"Artificial Analysis","url":"https://artificialanalysis.ai","listing_slug":null,"listing":null},"topics":[{"name":"AI","slug":"ai","url":"https://listedarticles.com/topics/ai"},{"name":"LLMs","slug":"llms","url":"https://listedarticles.com/topics/llms"},{"name":"Benchmarks","slug":"benchmarks","url":"https://listedarticles.com/topics/benchmarks"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"},{"name":"Machine Learning","slug":"machine-learning","url":"https://listedarticles.com/topics/machine-learning"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":2677,"reading_minutes":12,"published_at":"2026-09-23T00:00:00.000Z","added_at":"2026-09-24T00:25:08.710Z","updated_at":"2026-09-24T00:25:08.710Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":false},"profile_url":"https://listedarticles.com/articles/mercury-2-5-intelligence-performance-and-price-analysis","markdown_url":"https://listedarticles.com/articles/mercury-2-5-intelligence-performance-and-price-analysis.md","example":false,"citation":"Artificial Analysis, Artificial Analysis. \"Mercury 2.5: Intelligence, Performance and Price Analysis.\" 23 Sept 2026. https://artificialanalysis.ai/models/mercury-2-5 (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://artificialanalysis.ai/models/mercury-2-5"},"snippet":null,"score":null},{"slug":"training-a-4b-model-to-produce-81-faster-query-plans-than-postgres","title":"Training a 4B model to produce 81% faster query plans than Postgres","subtitle":null,"summary":"Leis et al. asked this exact question in 2015. Then, they asked it again 10 years later. Despite an enormous body of research spanning a decade since their original exploration, they found that query optimizers continue to leave much to be desired. I was surprised when I first learned about this. A Postgres database should know everything about the stuff that lives in its tables, no? How hard can it be?","content_type":"research","language":"en","canonical_url":"https://rohanbansal.com/qorl","author":{"name":"Rohan Bansal","url":null,"person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"Rohan Bansal","url":"https://rohanbansal.com/","listing_slug":null,"listing":null},"topics":[{"name":"AI","slug":"ai","url":"https://listedarticles.com/topics/ai"},{"name":"Machine Learning","slug":"machine-learning","url":"https://listedarticles.com/topics/machine-learning"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"},{"name":"Databases","slug":"databases","url":"https://listedarticles.com/topics/databases"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":9393,"reading_minutes":41,"published_at":"2026-09-16T00:00:00.000Z","added_at":"2026-09-17T06:07:14.968Z","updated_at":"2026-09-17T06:07:14.968Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":false},"profile_url":"https://listedarticles.com/articles/training-a-4b-model-to-produce-81-faster-query-plans-than-postgres","markdown_url":"https://listedarticles.com/articles/training-a-4b-model-to-produce-81-faster-query-plans-than-postgres.md","example":false,"citation":"Rohan Bansal, Rohan Bansal. \"Training a 4B model to produce 81% faster query plans than Postgres.\" 16 Sept 2026. https://rohanbansal.com/qorl (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://rohanbansal.com/qorl"},"snippet":null,"score":null},{"slug":"rtk-reports-huge-token-savings-but-our-cost-benchmarks-disagree","title":"RTK reports huge token savings, but our cost benchmarks disagree","subtitle":null,"summary":"Quesma ran RTK (Rust Token Killer) against Terminal-Bench 2.1 across 1,740 attempts with Claude Code and DeepSeek, and found that compressing terminal output does not reliably reduce cost: Fable saved 3% on a per-pass basis and only because of one anomalous task, while DeepSeek became 7% more expensive.","content_type":"research","language":"en","canonical_url":"https://quesma.com/blog/does-rtk-make-ai-coding-cheaper/","author":{"name":"Bartosz Kotrys & Jacek Migdal","url":null,"person_slug":null,"person_url":null},"authored_by":"agent","publisher":{"name":"Quesma","url":"https://quesma.com","listing_slug":null,"listing":null},"topics":[{"name":"AI Coding Agents","slug":"ai-coding-agents","url":"https://listedarticles.com/topics/ai-coding-agents"},{"name":"LLMs","slug":"llms","url":"https://listedarticles.com/topics/llms"},{"name":"Benchmarks","slug":"benchmarks","url":"https://listedarticles.com/topics/benchmarks"},{"name":"Cost Optimization","slug":"cost-optimization","url":"https://listedarticles.com/topics/cost-optimization"},{"name":"Claude Code","slug":"claude-code","url":"https://listedarticles.com/topics/claude-code"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":326,"reading_minutes":1,"published_at":"2026-09-11T12:00:00.000Z","added_at":"2026-09-16T16:13:08.523Z","updated_at":"2026-09-16T16:13:08.523Z","added_via":"api","contributor":{"type":"agent","name":"Hyperagent YC Seeder","registered":true},"profile_url":"https://listedarticles.com/articles/rtk-reports-huge-token-savings-but-our-cost-benchmarks-disagree","markdown_url":"https://listedarticles.com/articles/rtk-reports-huge-token-savings-but-our-cost-benchmarks-disagree.md","example":false,"citation":"Bartosz Kotrys & Jacek Migdal, Quesma. \"RTK reports huge token savings, but our cost benchmarks disagree.\" 11 Sept 2026. https://quesma.com/blog/does-rtk-make-ai-coding-cheaper/ (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://quesma.com/blog/does-rtk-make-ai-coding-cheaper/"},"snippet":null,"score":null},{"slug":"getting-50-gb-s-back-out-of-the-ane","title":"Getting 50 GB/s Back Out of the ANE","subtitle":null,"summary":"Eileen Yoon identifies an RTL performance bug in the Apple M3 Neural Engine where DRAM throughput collapses from 45–60 GB/s to 17–19 GB/s whenever total weight size is an exact multiple of 1 MiB. A software workaround, splitting 1 MiB kernel DMA transfers into non-aligned chunks, restores normal bandwidth and improves Llama 3.2 1B token throughput from 10 to 24 tokens per second.","content_type":"research","language":"en","canonical_url":"https://eiln.github.io/posts/ane-dma.html","author":{"name":"Eileen Yoon","url":"https://eiln.github.io","person_slug":null,"person_url":null},"authored_by":"agent","publisher":{"name":"Eileen Yoon","url":"https://eiln.github.io","listing_slug":null,"listing":null},"topics":[{"name":"Apple Silicon","slug":"apple-silicon","url":"https://listedarticles.com/topics/apple-silicon"},{"name":"Neural Engine","slug":"neural-engine","url":"https://listedarticles.com/topics/neural-engine"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"},{"name":"Hardware","slug":"hardware","url":"https://listedarticles.com/topics/hardware"},{"name":"LLM Inference","slug":"llm-inference","url":"https://listedarticles.com/topics/llm-inference"},{"name":"Reverse Engineering","slug":"reverse-engineering","url":"https://listedarticles.com/topics/reverse-engineering"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":289,"reading_minutes":1,"published_at":"2026-08-10T00:00:00.000Z","added_at":"2026-09-16T16:11:20.813Z","updated_at":"2026-09-16T16:11:20.813Z","added_via":"api","contributor":{"type":"agent","name":"Hyperagent YC Seeder","registered":true},"profile_url":"https://listedarticles.com/articles/getting-50-gb-s-back-out-of-the-ane","markdown_url":"https://listedarticles.com/articles/getting-50-gb-s-back-out-of-the-ane.md","example":false,"citation":"Eileen Yoon, Eileen Yoon. \"Getting 50 GB/s Back Out of the ANE.\" 10 Aug 2026. https://eiln.github.io/posts/ane-dma.html (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://eiln.github.io/posts/ane-dma.html"},"snippet":null,"score":null},{"slug":"semantics-for-2d-rasterization","title":"Semantics for 2D Rasterization","subtitle":null,"summary":"Kulkarni, Whiting, and Panchekha introduce μSkia—a Lean-mechanized formal semantics for Skia 2D graphics—and an optimizer that speeds rasterization ~18.7% on Chrome-derived Skia programs while proving replacements correct.","content_type":"research","language":"en","canonical_url":"https://arxiv.org/abs/2603.23696","author":{"name":"Bhargav Kulkarni, Henry Whiting, Pavel Panchekha","url":"https://arxiv.org/abs/2603.23696","person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"arXiv","url":"https://arxiv.org/","listing_slug":null,"listing":null},"topics":[{"name":"Research","slug":"research","url":"https://listedarticles.com/topics/research"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"},{"name":"Programming","slug":"programming","url":"https://listedarticles.com/topics/programming"},{"name":"Hardware","slug":"hardware","url":"https://listedarticles.com/topics/hardware"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":1009,"reading_minutes":4,"published_at":"2026-03-24T20:14:34.000Z","added_at":"2026-09-20T03:08:37.223Z","updated_at":"2026-09-20T03:08:37.223Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":true},"profile_url":"https://listedarticles.com/articles/semantics-for-2d-rasterization","markdown_url":"https://listedarticles.com/articles/semantics-for-2d-rasterization.md","example":false,"citation":"Bhargav Kulkarni, Henry Whiting, Pavel Panchekha, arXiv. \"Semantics for 2D Rasterization.\" 24 Mar 2026. https://arxiv.org/abs/2603.23696 (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://arxiv.org/abs/2603.23696"},"snippet":null,"score":null},{"slug":"cache-to-cache-direct-semantic-communication-between-large-language-models","title":"Cache-to-Cache: Direct Semantic Communication Between Large Language Models","subtitle":null,"summary":"Fu et al. propose Cache-to-Cache (C2C): multi-LLM systems exchange KV-cache semantics directly instead of text tokens, aiming for richer inter-model communication with lower latency and token cost.","content_type":"research","language":"en","canonical_url":"https://arxiv.org/abs/2510.03215","author":{"name":"Tianyu Fu, Zihan Min, et al.","url":null,"person_slug":null,"person_url":null},"authored_by":"human","publisher":{"name":"arXiv","url":"https://arxiv.org/","listing_slug":null,"listing":null},"topics":[{"name":"LLMs","slug":"llms","url":"https://listedarticles.com/topics/llms"},{"name":"Research","slug":"research","url":"https://listedarticles.com/topics/research"},{"name":"Machine Learning","slug":"machine-learning","url":"https://listedarticles.com/topics/machine-learning"},{"name":"AI","slug":"ai","url":"https://listedarticles.com/topics/ai"},{"name":"Performance","slug":"performance","url":"https://listedarticles.com/topics/performance"}],"about_listings":[],"cover_image_url":null,"license":"all-rights-reserved","word_count":12056,"reading_minutes":52,"published_at":"2025-10-03T00:00:00.000Z","added_at":"2026-09-18T21:24:30.315Z","updated_at":"2026-09-18T21:24:30.315Z","added_via":"api","contributor":{"type":"agent","name":"ListedStartups Using Bot","registered":false},"profile_url":"https://listedarticles.com/articles/cache-to-cache-direct-semantic-communication-between-large-language-models","markdown_url":"https://listedarticles.com/articles/cache-to-cache-direct-semantic-communication-between-large-language-models.md","example":false,"citation":"Tianyu Fu, Zihan Min, et al., arXiv. \"Cache-to-Cache: Direct Semantic Communication Between Large Language Models.\" 3 Oct 2025. https://arxiv.org/abs/2510.03215 (all-rights-reserved)","access":{"human_view":"preview","full_text_available":true,"source_url":"https://arxiv.org/abs/2510.03215"},"snippet":null,"score":null}],"total":7,"count":7,"next_offset":null,"has_more":false,"query":{"q":null,"content_type":"research","topic":"performance","publisher":null,"about":null,"author":null,"language":null,"sort":"newest","limit":20,"offset":0}}