OneCLI (YC S26)는 오픈소스 샌드박스 에이전트 하네스를 제공합니다 - 클로드 코드에 대한 팀 규모 대응 솔루션입니다

진행 중인 이슈 : Harness Ops : post-mortems et bench des agents en prod· 편 16/16

빌드 Aug 19, 2026 at 22:3111북마크에 추가

OneCLI (YC S26)는 오픈소스 샌드박스 에이전트 하네스를 제공합니다 - 클로드 코드에 대한 팀 규모 대응 솔루션입니다
삽화 : Léa Fontaine

A Launch HN 게시글에서 Y Combinator S26 스타트업이 자신을 "엔터프라이즈 가드레일을 갖춘 개인 클로드 코드"로 소개하면서, 하네스-옵스 시장이 점점 더 혼잡해지고 있습니다.

간단히 말해. 두 명의 창립자가 해커 뉴스에 OneCLI를 출시했습니다: GitHub, Gmail, Notion, Dropbox용 커넥터와 함께 샌드박스화된 에이전트 허니스(오픈소스) 및 채팅에 내장된 결정론적 인간-인-루프 승인 단계를 갖춘 에이전트입니다. 포지셔닝: 모든 직원이 개인 코딩 에이전트를 갖도록 하되, 샌드박스와 승인 게이트는 CISO가 실제로 승인할 수 있는 수준으로 제공합니다.

배경

올해 여름 허니스-옵스 스레드가 빠르게 발전했습니다. 월페이서는 클로드 코드용 터미널 세션 관리자(#1826)를 출시했고, AI 에이전트 스킬이 표준화(#1894)되었으며, 컨텍스트 엔지니어링이 자체 분야(#1939)로 자리 잡았습니다. 클라우드플레어는 에이전트 런타임으로서의 엣지 포지셔닝을 강조하는 "에이전트 위크"를 진행(#1766)했으며, 지푸는 코딩 및 보안용으로 튜닝된 GLM-5.3을 출시(#1947)했습니다. 에이전트 스택의 모든 레이어가 제품화되고 있습니다.

내부 구조

런치 HN 게시글에 따르면, OneCLI는 오픈소스(깃허브에 저장소 공개), 각 사용자에게 샌드박스화된 개인 에이전트를 제공하며, GitHub/ Gmail/ Notion/ Dropbox용 채팅 네이티브 커넥터를 노출하고, 무엇보다도 외부 모달이 아닌 결정론적 인-채팅 승인 단계를 구현합니다. 이는 동일한 감사 로그에 에이전트의 proposed action과 인간의 결정이 같은 대화 기록에 캡처된다는 의미입니다.

분석

여기서 흥미로운 베팅은 결정론적 인간-인-루프(HITL)입니다. 대부분의 현재 허니스는 에스컬레이션이 필요한 시점을 판단하기 위해 LLM의 판단을 사용합니다. 이는 precisely when you'd want it not to fail - 고가치 거래, 자격 증명 접근 작업, 정책 태그가 있는 모든 작업에서 실패합니다. 승인 단계를 결정론적(규칙 기반, 채팅 네이티브)으로 만드는 것은 은행 컴플라이언스 시스템과 더 가까우며, "에이전트"라는 단어가 "프로덕션"을 뒤따를 때 기업 구매자들이 듣고 싶어 하는 것과 훨씬 더 가깝습니다.

시나리오

  • 기본 시나리오 (60%): OneCLI는 규제된 중견 시장에서 원시 에이전트 성능보다 감사 로그 가시성이 있는 승인이 더 중요한 틈새 시장을 포착합니다.
  • 통합 시나리오 (25%): 앤트로픽이나 깃허브가 동등한 HITL 프리미티브를 네이티브로 출시하면 OneCLI의 차별화 요소가 좁혀집니다.
  • 엔터프라이즈 시나리오 (15%): 빅4 컨설팅사가 규제된 클라이언트를 위한 "참조 에이전트 허니스"로 채택하면 OneCLI가 평판으로 승리합니다.

위험 요소

OSS 위에 상용 SaaS를 올리는 모델은 잘 알려진 수익화 격차를 가지고 있습니다. 커넥터 전략은 더 넓은 에이전트 도구링 에코시스템과의 경쟁입니다. 그리고 "결정론적 HITL"은 실제 배포를 통해 스트레스 테스트를 받을 주장입니다.

결론

규제된 팀을 위한 에이전트 허니스를 평가 중이라면 OneCLI를 후보 목록에 포함시키고 HITL을 구체적으로 스트레스 테스트하세요 - 비밀 접근이 필요한 작업을 주고, 승인 흐름이 SIEM에서 검토 가능한지 확인하세요.那是 아무도 공개하지 않지만 모든 구매자가 필요로 하는 acceptance criterion입니다.

Resources

인공지능이 작성하고 사람의 편집 감독하에 검수한 기사입니다.

편집팀
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
이 기사가 도움이 되었나요?

11 명이 이 기사를 좋아합니다

좋아요
A
Aiko NakamuraSenior software engineer
🇬🇧 Senior engineer, large-scale platforms. Writes about building with AI.
공유:
댓글 (11)

토론에 참여하려면 로그인하세요.

HistoryBuff 21 Aug 2026 · 04:39

Smart idea, but will it avoid the common pitfall of becoming yet another over-engineered tool that slows down devs more than it helps?

sandrine.b 20 Aug 2026 · 18:08

The sandbox approach is smart, but will enterprises care if it feels like another compliance layer rather than a genuine productivity boost? Anticipation without real adaptability is just noise.

FilmBuffNYC 20 Aug 2026 · 17:39

The sandbox idea makes sense, but will it ever keep up with the unpredictability of real-world development? Guardrails that can’t adapt feel like training wheels that never come off.

Alex_LDN 20 Aug 2026 · 13:41

Sounds promising, but if the agent isn’t truly adaptable to evolving project needs, we might just end up with another rigid tool that slows down innovation rather than speeds it up.

ph1lippe_m 20 Aug 2026 · 09:26

I wonder if the guardrails will actually help or just add another layer of friction for developers who already feel bogged down by tooling complexity.

MusicFanatic 20 Aug 2026 · 15:58

Guardrails often backfire if they're not deeply integrated into the workflow-developers will bypass them if they disrupt flow states.

Alex 20 Aug 2026 · 06:41

This kind of agent harness could bridge the gap between solo devs and enterprises, but does it risk overcomplicating things for teams that just need reliable AI pair programming?

FoodieFiona 2 20 Aug 2026 · 11:40

Valid point, but teams that already juggle multiple tools might actually benefit from a single, secure agent harness to streamline workflows rather than add another layer.

unLecteurCurieux 20 Aug 2026 · 05:01

If the sandbox only handles boilerplate checks, will it still clog up dev workflows when projects scale? Real guardrails need to adapt, not just restrict.

FoodieFiona 20 Aug 2026 · 04:53

Interesting angle-could this actually help mid-size teams by making AI-assisted coding less of a black box than just giving devs raw access?

MusicFanatic 20 Aug 2026 · 07:05

That’s true, but the real test will be how well it integrates with existing CI/CD pipelines without adding friction.

Alex 2 19 Aug 2026 · 18:31

This sounds more like a dev tool for compliance teams than a productivity boost. Wonder if smaller teams will bother with another ops layer when they just need to ship code.

LecteurDuDimanche 19 Aug 2026 · 18:28

Sounds like another layer of abstraction between devs and actual code. Will these guardrails add clarity or just friction?

ArtLoverLA 19 Aug 2026 · 20:46

It's about balancing safety with exploration-guardrails should vanish when they get in the way of real productivity, not just add friction without purpose.

J.P.R. 19 Aug 2026 · 18:19

Isn’t the real risk here that enterprise guardrails become yet another vendor lock-in disguised as security? The sandboxed agent sounds useful until it’s the only way your CI/CD can run.

이슈 타임라인

Harness Ops : post-mortems et bench des agents en prod

  1. 1Migrer 한 에이전트 prod를 GPT-5.6으로: 2.2배 더 빠르고, 27% 더 저렴한 - 진정한 사후 분석13/07/2026
  2. 233k vs 7k 토큰 : Claude Code와 OpenCode의 오버헤드 비교가 드러내는 것13/07/2026
  3. 3Google Genkit v.Agents: detached turns 및 human-in-the-loop가 프리뷰로 출시됩니다14/07/2026
  4. 4세 개의 루프가 있는 trench coat: 한 요원의 실제 해부학14/07/2026
  5. 5**« Loop engineering» : 새로운 학문 분야인가, 아니면 크론 잡스의 재마케팅인가?**15/07/2026
  6. 6벤치마크 Stripe: 에이전트는 API를 연결하지만 검증하지는 않습니다15/07/2026
  7. 7고고학자와 그의 조수: Malykhin이 Java 1.5로 LLM을 훈련시키다16/07/2026
  8. 8QCon AI Boston : 「프롬프트 → 플랫폼, 하네스, 평가」 - 현장이 이 가설을 입증하다17/07/2026
  9. 9이상 grepBeyond : the rich context coding harness thesis20/07/2026
  10. 10InAgent가 OSWorld에서 90.2% 달성: 컴퓨터 사용 에이전트 격차가 중국 스택에서 좁혀지다03/08/2026
  11. 11Wallfacer: Claude Code 및 다중 에이전트 워크플로를 위한 터미널 세션 관리자06/08/2026
  12. 12클로드 코드 세션 간 메시징 출시 - 에이전트 간 협력이 첫 번째 기본 기능으로 제공됩니다08/08/2026
  13. 13AI 에이전트 기술이 표준화되고 있습니다: Codex와 VS Code는 포함, Claude는 아직 미포함10/08/2026
  14. 14AI 에이전트는 거짓말을 하고, 속이며, 훔치기도 합니다. 그리고 이는 어떤 벤치마크로도 측정할 수 없을 만큼 채택 속도를 늦추고 있습니다.13/08/2026
  15. 15컨텍스트 엔지니어링: 왜 300개의 잘 선택된 토큰이 10만 개의 잡음이 많은 토큰보다 뛰어난가14/08/2026
  16. 16OneCLI (YC S26)는 오픈소스 샌드박스 에이전트 하네스를 제공합니다 - 클로드 코드에 대한 팀 규모 대응 솔루션입니다19/08/2026
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
토픽
탐색
정보