PublishersPoolside
Poolside
5 documents read from this publisher.
poolside/Laguna-S-2.1 · Hugging Face
first party release · model card · 2026-08-03 · 9 techniques
Model release notes - Poolside
first party release · vendor docs · 2026-07-22 · 13 techniques
Mixture of ExpertsConfigurable reasoning effortFP8 KV-cache quantizationSparse expert activationInterleaved thinking between tool callsAgentic edit-and-validate loop (read, edit, run checks)Checking for agent-generated test scripts before committingCheckpointingIncremental task decomposition for agent promptsInstruction-following fine-tuningPlanning with the model for self-contained tasksReinforcement Learning from Code Execution FeedbackTask-specification prompting
Introducing Laguna S 2.1
first party release · official blog · 2026-04-28 · 23 techniques
Mixture of Expertsavg@kFP8-precision reinforcement learningReinforcement learning for low-pass-rate tasksSFT bootstrapping with synthetic dataagentic repository installation taskautomated agentic research loop for software optimizationbenchmark-instrumented target code for agent optimizationFull trajectory releaseHuman-calibrated LLM-as-a-JudgeLong-horizon rollout budgetsModel FactoryMulti-harness rolloutsone-change-at-a-time benchmarking with measurable-wins retentionPrompt Addendum Prohibiting Direct Online-Solution UseRL sandboxing with selective network blocking and artifact cachingsandboxed agentic benchmarking on a Harbor fork with step limitSeed-based terminal task generationThinking mode with automatic test-time compute budgettraining objective for persistence and verification behaviorsTraining-task generation from real commit historyUser-configurable thinking-effort controlXML-tagged tool-call format
Two foundation models built for agentic coding.
first party release · vendor docs · 2025-10-28 · 2 techniques
laguna m1 xs2 technical report
technical report · 115 techniques
Mixture of ExpertsHybrid AttentionFP8 KV-cache quantizationGrouped-query attentionMuonShared ExpertsRotary Position EmbeddingWarmup-Stable-DecayContinual pretraining for long-context extensionCross-turn persistent reasoning historyToken-in-token-out (TITO)512-token Sliding Window AttentionAutoMixerRemoving Git-History Leaks from Benchmark ImagesAgentic RL Task MixAgentic WorkflowsAtlas inference libraryAutomatic in-flight checkpoint evaluation schedulingAutomating Repetitive Infrastructure WorkAuxiliary-loss load balancingAWQ INT4 weight quantization (W4A16)BF16 mixed-precision trainingBinary task verifierBinary terminal-verifier rewardBlind pairwise bucket-boundary calibrationBlocking in-flight rollout steps on weight updatesCapping trajectory stalenessCISPO with Length-Weighted Leave-One-Out Group-Relative AdvantagesClosed-loop multi-turn rolloutConfiguration-flag promotion of validated components to productionConservative model-based noise filteringContinuous contribution-score rankingCross-replica model-weight hash consistency checksCurrent-turn-only reasoning-mode detectionDAG-based asset lineage trackingDecompositional instruction-following judgeDense annotation of ambiguous low-quality dataDense gated attention with full RoPE and full gatingDeterministic chain of checkersDiverse benchmark validation for quantization