LingBot-Map은 streaming 3D를 ‘긴 컨텍스트’가 아니라 역할이 다른 기하 메모리로 푼다
LingBot-Map은 단안 이미지 스트림에서 camera pose와 depth를 계속 추정하는 feed-forward 3D foundation model이다. anchor·pose-reference window·trajectory memory를 분리한 Geometri...
Blog
LingBot-Map은 단안 이미지 스트림에서 camera pose와 depth를 계속 추정하는 feed-forward 3D foundation model이다. anchor·pose-reference window·trajectory memory를 분리한 Geometri...
Cua는 기존 desktop을 제어하는 Driver, 격리된 computer를 만드는 Sandbox·Lume·Fleet, 그리고 task와 evaluator를 분리하는 Cua-Bench를 묶어 computer-u...
Dream-RSI는 discovery tree에 남은 실행 outcome을 replay simulator로 재해석해, exploration policy 후보를 실제 rollout 없이 비교하고 다음 온라인 탐색으...
jina-ocr-v1은 DeepSeek-OCR의 압축형 vision encoder와 3B MoE decoder 위에 FastMTP speculative decoding과 dense verifiable reward...
TypeSafe AI의 Jev는 자유 텍스트를 생성하는 대신, 코드가 정의한 Noul·Choice·Score 질문에 타입화된 값·확률·신뢰도를 반환하도록 설계된 System One 모델이다.
COBRA-Skills는 contextual bandit으로 실행 평가 예산을 유망하거나 덜 탐색된 agent skill에 배분하고, 실제 rollout 증거로 후보군을 재생성·변이·교차하는 skill-optim...
NeoHorse-1은 에이전트 실행 중 남는 라우팅·도구·결과 기록을 커리큘럼과 온폴리시 증류, 다음 데이터 배분에 연결해 4B·9B 오픈 가중치 모델을 post-training한 초기 RSI 프로토타입이다.
Qwen-Drive-1.0은 Qwen3.5-4B의 시각·언어 경로는 유지한 채 외부 BEV 인지 헤드와 궤적 생성 Planning Expert를 붙여, 3D 인지·주행 VQA·모션 플래닝을 하나의 공개 패키지로...
Repo-To-Skill은 DisCo가 repository·paper·task 자료에서 절차·검증·복구 경로를 skill graph로 증류하고, AREX-Skill Library가 필요한 분기만 research...
H Company의 NeoMME는 텍스트 토큰과 원본 이미지 패치를 하나의 양방향 Transformer로 처리하고, dense·late-interaction 검색 표현을 한 번에 내보내며 시각 문서 RAG의 모델...
Harness-of-Harness는 planner·developer·QA tester를 artifact와 evidence로 연결해, coding agent가 여러 iteration에 걸쳐 계획·구현·검증을 누적하...
Skaling은 모델 크기와 학습 토큰 수가 독립적으로 loss를 낮춘다는 Chinchilla의 가정을 하나의 결합 지수로 완화하고, 저비용 L-shape 격자만으로 대규모 학습 구간의 loss를 예측하려는 스케...
WikiSkill은 실행 trace·지속적으로 누적되는 wiki·실행 가능한 skill을 분리하고, 실패·성공 경험을 패턴과 수용 이력으로 축적한 뒤 validation gate를 통과한 skill 변경만 반영하...