反馈类型
详细描述
Please improve the context-management feature and documentation in ETOS LLM Studio. The current UI exposes Context Window Count and Lazy Loading Rounds, but their relationship to model context limits and context compression is unclear. Please consider: (1) clearly documenting the difference between Context Window Count, Lazy Loading Rounds, model context length, and maximum output tokens; (2) adding an explicit automatic context-compression/summarization option, with configurable trigger threshold and retention policy; (3) indicating whether older messages are summarized, truncated, or merely loaded on demand; (4) displaying current context/token usage and warning before the model context limit is reached; (5) providing recommended values for different model types; and (6) clarifying the performance and memory implications of setting both values to 0.
可复现步骤
- Open ETOS LLM Studio 1.7.0 (Build 372) on iOS. 2. Open Settings. 3. Open Preferences. 4. Scroll to Context Window Management. 5. Observe Context Window Count and Lazy Loading Rounds, both set to 0, and the brief explanatory text. 6. Try to determine how automatic context compression or summarization is configured.
预期行为
Users should be able to understand and configure how long conversations are loaded and compressed, know when compression or truncation occurs, and see clear warnings or usage information before reaching the model's context limit.
实际行为
In ETOS LLM Studio 1.7.0 (Build 372), Settings > Preferences > Context Window Management shows only two options: Context Window Count and Lazy Loading Rounds. Both are currently set to 0. The description says that 0 loads all conversation history, but the UI does not explain whether these settings are related to token limits, automatic summarization, truncation, or context compression. I could not find a clear, dedicated context-compression setting or documentation in the app.
补充信息
App: ETOS LLM Studio; Version: 1.7.0 (Build 372); Git commit shown in About: c4090fd; Developer: Eric-Terminal; Platform: iOS/watchOS. Relevant UI path shown in screenshots: Settings > Preferences > Context Window Management. Current values: Context Window Count = 0; Lazy Loading Rounds = 0. The About page links to the project at github.com/Eric-Terminal/ETOS-LLM-Studio and documentation at docs.els.ericterminal.com. This is a feature/documentation suggestion rather than a report of a crash.
环境信息
- 平台: ios
- App 版本: 1.7.0 (Build 372)
- Git 提交: c4090fd
- 分发通道: App Store
- 系统版本: Version 26.5 (Build 23F77)
- 设备型号: iPhone17,2
- 语言: zh-Hans_HK
- 时区: Asia/Hong_Kong
最小诊断日志
- timestamp=2026-07-13T04:11:48Z
- provider_count=2
- session_count=10
- platform=iOS
- distribution_channel=appStore
- app_version=1.7.0(372)
服务端附注
- 来源: source/app-feedback
- 同步标记: 由用户提出自动更新的
- 客户端IP哈希: 1484711a5ebe730a77401e48ce33909abaa282b4b17dbb6c2a61b5c28bd37308
反馈类型
详细描述
Please improve the context-management feature and documentation in ETOS LLM Studio. The current UI exposes Context Window Count and Lazy Loading Rounds, but their relationship to model context limits and context compression is unclear. Please consider: (1) clearly documenting the difference between Context Window Count, Lazy Loading Rounds, model context length, and maximum output tokens; (2) adding an explicit automatic context-compression/summarization option, with configurable trigger threshold and retention policy; (3) indicating whether older messages are summarized, truncated, or merely loaded on demand; (4) displaying current context/token usage and warning before the model context limit is reached; (5) providing recommended values for different model types; and (6) clarifying the performance and memory implications of setting both values to 0.
可复现步骤
预期行为
Users should be able to understand and configure how long conversations are loaded and compressed, know when compression or truncation occurs, and see clear warnings or usage information before reaching the model's context limit.
实际行为
In ETOS LLM Studio 1.7.0 (Build 372), Settings > Preferences > Context Window Management shows only two options: Context Window Count and Lazy Loading Rounds. Both are currently set to 0. The description says that 0 loads all conversation history, but the UI does not explain whether these settings are related to token limits, automatic summarization, truncation, or context compression. I could not find a clear, dedicated context-compression setting or documentation in the app.
补充信息
App: ETOS LLM Studio; Version: 1.7.0 (Build 372); Git commit shown in About: c4090fd; Developer: Eric-Terminal; Platform: iOS/watchOS. Relevant UI path shown in screenshots: Settings > Preferences > Context Window Management. Current values: Context Window Count = 0; Lazy Loading Rounds = 0. The About page links to the project at github.com/Eric-Terminal/ETOS-LLM-Studio and documentation at docs.els.ericterminal.com. This is a feature/documentation suggestion rather than a report of a crash.
环境信息
最小诊断日志
服务端附注