ds2api

mirror of https://github.com/CJackHwang/ds2api.git synced 2026-05-06 01:15:29 +08:00

Author	SHA1	Message	Date
CJACK	112bedb05d	refactor: differentiate reference marker handling between stream and non-stream modes - Stream: strip both and [reference:N] markers to prevent leaking partial link metadata during incremental output - Non-stream: convert citation/reference markers to Markdown links for Claude Messages, Gemini generateContent, and OpenAI Chat/Responses - Remove StripReferenceMarkers option from call sites; behavior is now determined automatically by stream vs non-stream context - Extend JS runtime stripReferenceMarkersText() to also match [citation:N] - Add tests for streaming marker stripping and non-stream link conversion Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-03 17:53:49 +08:00
CJACK	c099a6f7bf	feat: add unified response history session management across Claude, Gemini, and OpenAI API backends Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-05-03 17:24:38 +08:00
CJACK	1286b02247	refactor: remove legacy compatibility configuration and UI components	2026-05-03 04:14:19 +08:00
CJACK	5f110e6910	refactor: remove legacy history split configuration and integrate current input file handling into the completion runtime pipeline.	2026-05-03 01:50:50 +08:00
CJACK	dc5bffdf89	refactor: centralize assistant turn semantics and stream accumulation into new assistantturn and completionruntime packages	2026-05-02 23:28:43 +08:00
CJACK.	7c3ff6ee7e	Merge pull request #374 from shern-point/feat/full-context-file-token-accounting Feat/full context file token accounting	2026-04-30 02:12:55 +08:00
CJACK.	63e62fd1b0	Merge pull request #372 from shern-point/feat/accurate-context-token-length Feat/accurate context token length	2026-04-30 02:11:32 +08:00
shern-point	6a778e0d35	feat: include inline-uploaded file tokens in context token accounting Track byte sizes of inline-uploaded files during PreprocessInlineFileInputs and convert them to conservative token estimates (bytes/3). RefFileTokens is threaded through StandardRequest into all OpenAI chat/responses usage builders so returned prompt_tokens/input_tokens reflect the full upstream context cost including attached files.	2026-04-30 01:42:51 +08:00
shern-point	415a2359ad	feat: route OpenAI responses usage through preserved prompt text Use the stored full-context prompt text for responses accounting so neutral placeholder prompts do not underreport returned input token counts.	2026-04-30 00:45:31 +08:00
CJACK.	33f6fef015	Fix tool-call fallback on sanitized empty text and remove history wrapper tags	2026-04-29 23:04:45 +08:00
MiY	241334c658	Fix stream compatibility and vision model exposure	2026-04-29 20:23:13 +08:00
shern-point	b9c8e90d98	refactor: thread tool schemas through responses tool outputs Ultraworked with [Sisyphus](https://github.com/code-yeongyu/oh-my-openagent) Co-authored-by: Sisyphus <clio-agent@sisyphuslabs.ai>	2026-04-28 13:46:06 +08:00
CJACK	28bb85ad63	refactor: replace history_split with current_input_file configuration Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-27 23:36:56 +08:00
CJACK	0378d8c0a9	feat: add empty-output retry and Vercel auto-continue support - Auto-retry Chat/Responses streams once when upstream output is empty but not content-filtered, reusing session/token/PoW and appending a regeneration suffix to the prompt - Wire DeepSeek continue API into Vercel streams for multi-round thinking output exhaustion - Defer empty-output errors in stream finalizers to enable synthetic retry; only surface failure when the retry budget is exhausted - Track content_filter stops to avoid retry on filtered outputs - Add comprehensive tests for stream/non-stream retry, Responses retry, and content_filter no-retry - Update prompt-compatibility.md documentation Co-Authored-By: Claude Opus 4.7 <noreply@anthropic.com>	2026-04-27 18:00:52 +08:00
CJACK	40d5e3ebb5	测试DSML	2026-04-27 00:21:26 +08:00
MiY	a505f2cb96	fix: fallback tool calls from thinking on empty output	2026-04-26 17:45:12 +08:00
CJACK	abc96a37d8	refactor backend API structure	2026-04-26 06:58:20 +08:00

17 Commits