Commit graph

21 commits

Author SHA1 Message Date
0bfcb0f189 feat: usage metrics - identity metering, usage token logging, and runtime proto 2026-07-10 17:54:23 +09:00
toki
ff7330680f refactor: reject transformed mode and responses for provider-pool/tunnel routes 2026-07-09 12:47:20 +09:00
toki
b6363f7dd2 feat(openai): explicit reasoning 보존 로직 추가 및 관련 핸들러 개선
Strict output 모드에서 reasoning이 명시적으로 요청된 경우 이를 보존할 수 있도록
normalizeCompletionOutput 함수와 관련 핸들러들의 로직을 업데이트한다.
2026-07-09 11:23:34 +09:00
2b3793149c fix(edge/openai): remove preview fields from production logs
Remove all *_preview zap.String/Any fields from OpenAI-compatible log
calls (chat_handler, responses_handler, stream) so that prompt, content,
reasoning, source, and delta values are never written in plain to
operational logs.

Replace preview fields with non-content metadata:
  - prompt_len, content_len, reasoning_len, delta_len, source_len
  - message_count, content_len, finish_reason, tool_call_count

Also remove the now-unused previewString helper from types.go.

Add log_safety_test.go with regression tests that:
  - verify no *_preview fields exist on any log line
  - verify operational metadata fields are retained for observability
  - verify logOpenAICompatStreamOutput does not emit delta_preview
2026-07-06 15:21:50 +09:00
toki
ca5eb4685c fix(edge): provider-pool thinking과 queue 정책을 정리한다
vLLM-MLX provider-pool에서 strict output이 thinking을 끄지 않도록 하고, queue_timeout_ms=0을 IOP queue timeout 없음으로 해석해야 한다. long reasoning/long-context 요청은 IOP queue timeout이 아니라 caller cancellation과 backend timeout 정책으로 제어한다.
2026-07-06 10:16:43 +09:00
5b95208f4c feat: edge runtime model queue and status dispatch updates
- Update control-plane edge wire and edge-node runtime wire contracts
- Refactor model queue service with admission control
- Update chat handler and responses handler for edge
- Modify run dispatch and status provider logic
- Add/modify runtime proto definitions
- Move G07 status logs to archive
2026-07-05 20:01:27 +09:00
2c9faad1f3 feat: long context admission support - refactor edge server, add input estimator 2026-07-05 18:18:58 +09:00
77ab36cbd1 feat: edge config refresh and OpenAI handlers update
- Add config refresh classification and test coverage
- Update OpenAI chat/responses/stream handlers with config refresh support
- Add comprehensive server tests for config refresh
- Update local dev guide with config refresh information
- Extend config package with ConfigRefresh field
2026-07-04 09:33:00 +09:00
90bce8336f feat: edge openai handler updates and node provider first config surface
- Modify edge openai chat/stream handlers for priority routing
- Update responses_handler and run_result
- Add model queue test updates
- Add node provider first config surface task
- Archive inflight accounting recovery docs
2026-07-01 10:44:48 +09:00
ea7388cfe9 fix(openai): native tool unsupported fallback 추가 2026-06-27 22:24:06 +09:00
7dc0788f88 fix(openai): provider tool call을 그대로 전달한다 2026-06-27 21:44:49 +09:00
46bd8563e2 feat: OpenAI compatible API 변경 사항 적용
- agent-contract: OpenAI compatible API 계약 문서 업데이트
- apps/edge: OpenAI API smoke/test 스크립트 및 핸들러 개선
- scripts: e2e 테스트 스크립트 업데이트
2026-06-27 05:16:25 +09:00
leedongmyung
0327d11dcc update: dev-corp 환경 추가 2026-06-25 15:15:56 +09:00
06f1d36fa3 feat: provider catalog device status - edge dispatch implementation
- Add agent-task archive for 02+01_edge_dispatch
- Remove outdated plan and code review docs
- Update edge bootstrap, input manager, and openai handlers
- Update model queue, run dispatch, and service layer
- Add status provider tests
2026-06-20 16:59:55 +09:00
747fa27fa7 feat: update inference provider extension phase and edge app changes
- Update PHASE.md for inference-provider-extension
- Archive vllm-provider-serving-validation milestone and SDD
- Update edge app OpenAI handlers (chat, responses, stream, run_result)
- Update tests and e2e scripts for vLLM integration
- Update READMEs for edge and node apps
2026-06-18 09:45:38 +09:00
957700c2d6 feat: model route queue policy alignment phase added and edge service updates
- Add model-route-queue-policy-alignment milestone to inference provider extension phase
- Add SDD documentation for inference provider extension
- Update edge chat/responses handlers for OpenAI compatible API
- Update edge config with new queue policy settings
- Add config tests for queue policy support
- Update task tracking for model route queue policy alignment
2026-06-16 22:30:36 +09:00
dba289d0fa feat(edge): 모델 그룹 큐 스케줄링 구현
- 모델 그룹별 큐 기반 디스패팅 로직 추가
- ModelQueue 관리자로 모델 인스턴스 큐 처리
- OpenAPI chat 및 responses 핸들러 업데이트
- edge 서비스 테스트 코드 개선
2026-06-16 10:16:05 +09:00
c34fec8f6b chore: update edge openai handlers and reorganize task archive 2026-06-13 17:18:21 +09:00
a758926593 feat: openai workspace agent execution contract milestone and edge updates
- Add milestone doc for openai-workspace-agent-execution-contract
- Update edge chat/responses handlers and types
- Update edge service run dispatch logic
- Add edge and service tests
- Update edge README
2026-06-13 16:33:00 +09:00
738bd1939c feat: m-node-multi-target-serving-foundation G07/G08 work - route catalog, engine profiles, edge/node updates 2026-06-11 14:43:24 +09:00
80fcbca74f feat(edge): Responses 입력 표면을 추가한다
NomadCode Core가 OpenAI Responses shape로 Edge 실행을 요청할 수 있도록 non-streaming /v1/responses 경계와 metadata 전달, smoke 검증을 함께 고정한다.
2026-06-08 15:16:25 +09:00