Research Digest — 2026-08-21
17 entries today · 3 apply · 9 shelf · 5 ignored · 14 papers
Apply Now
- ATH-MaaS/Pixelle-Video — other · video-generation, short-form, automation, composable, python, multimodal → AI Learning Path (tool)
- docling-project/docling — data-pipeline · document-parsing, pdf-extraction, data-pipeline, multiformat, xbrl, agent-ready → Auto Social Posting, ClientCo Intelligence Line (agentic research) (tool)
- stablyai/orca — dev-tooling · agent-orchestration, worktrees, claude-code, desktop, mobile, multi-agent → Research Console (this tool), ClientCo Intelligence Line (agentic research) (project)
Research Papers
- Tuning the Stochastic Machine: A Systems Engineer's Operating Model for Human-AI Engineering · prompting · adapt — Systematically persist LLM error corrections across sessions to prevent repeat failures.
- Eureka: Task-Conditioned Meta-Agent Orchestration for Scientific Discovery · agents · adapt — Orchestrate specialized agents dynamically via task-specific obligation graphs and memory.
- Self-prompting and cross-model consensus enable reproducible data extraction from scientific literature with large language models · prompting · adapt — Extract consistent data from papers via self-prompting and multi-model consensus.
- Harness Continual Learning: Continual Adaptation Beyond Model Parameters · memory · adapt — Agents adapt through prompts, memories, tools, skills without retraining models.
- DeepWeaver: Bridging the Evidence Synthesis Gap in Open-Ended Question Answering · rag · adapt — Organize fragmented RAG evidence into well-cited comprehensive answers via synthesis.
- Beyond Teacher Likelihood: Group-Calibrated On-Policy Distillation for Long-Context Reasoning · training · read — Dense token-level guidance from teachers helps students avoid locally-plausible incomplete reasoning.
- Beyond the Transcript: Detecting Covert Co ordination in Latent Multi-Agent Communication · agents · read — Detect hidden coordination between LLM agents via their internal activation patterns.
- When Readability and Source Retention Diverge: An Evaluability Gap in AI Translation · evals · read — AI translations read well but may not preserve source meaning under evaluation.
- Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk · other · read — Model financial risk under evolving uncertainty via robust optimization identity.
- What is Missing from AI Post-Training AI: An Empirical Analysis · agents · read — AI agents struggle with end-to-end training; conflates training execution with post-training.
- Adaptive Memory and Reflection Multi-Agent System for Medical Question Answering · memory · read — Multi-agent medical QA improves by adapting memory and reflecting on evidence.
- Grading the Graders: Verification Autonomy Levels (L0-L5) for LLM Reasoning · evals · read — Standardize verifier automation levels (L0-L5) for LLM reasoning validation.
- A Theory of Post-hoc Debate Judgement · reasoning · read — Debate between agents improves reasoning and explainability via internal discussion.
- rEDMRec: Distilling Large Language Model Reasoning into an Editable Experience Memory for Recommendation · memory · read — Improve recommendations by capturing LLM reasoning over history in editable memory.
Future Shelf
- Osmantic/ODS — llm-runtime · local-ai, self-hosted, workflows, agents, rag, inference — WATCH
- Tencent/AI-Infra-Guard — security · agent-scanning, mcp-security, skill-validation, red-teaming, ai-governance, jailbreak-detection — WATCH
- agent-substrate/substrate — llm-runtime · agent-infrastructure, kubernetes, multiplexing, sandbox, google, agentic-os — WATCH
- microsoft/agent-framework — agent-framework · orchestration, multi-agent, python-dotnet, observability, workflow, governance — WATCH
- modular/modular — llm-runtime · mojo, max-framework, inference, ai-platform, compiler, accelerator — WATCH
- pipecat-ai/pipecat — agent-framework · voice-agents, multimodal, real-time, orchestration, python, multi-agent — WATCH
- smart-mcp-proxy/mcpproxy-go — dev-tooling · mcp, proxy, ai-agents, security, tool-federation, cross-platform — WATCH
- smtg-ai/claude-squad — claude-tooling · session-manager, worktree-isolation, tmux, multi-agent, git, tui — WATCH
- sourcebot-dev/sourcebot — dev-tooling · code-search, codebase-analysis, self-hosted, llm-powered, code-navigation, docker — WATCH
Ignored
- AprilNEA/OpenLogi — dev-tooling · logitech, device-config, rust, privacy, cli-gui, cross-platform
- goauthentik/authentik — security · sso, idp, authentication, oauth2, self-hosted, identity
- mahlernim/google-timeline-visualizer — other · travel, data-visualization, location-history, android, video-export, maps
- makeplane/plane — dev-tooling · project-management, issue-tracking, open-source, self-hosted, sprint-planning, team-collaboration
- sonnyflylock/voxie-ai-directory-mcp — claude-tooling · mcp-server, voxie-ai, directory-service, phone-number, webchat, api-wrapper