fleet-memory/hindsight-api/hindsight_api/engine
Chris Bartholomew 49ae55af03
Switch Vertex AI provider to native genai SDK (#242)
Replace the OpenAI-compatible endpoint approach with the native
google-genai SDK for Vertex AI. This eliminates the custom token
refresher, TokenInjectingTransport, and async lifecycle complexity
while also removing the 8192 output token cap that the OpenAI
endpoint enforced.

Changes:
- vertexai provider now uses genai.Client(vertexai=True) instead of
  AsyncOpenAI with token-injecting transport
- Routes through existing _call_gemini/_call_with_tools_gemini paths
- Strips google/ prefix from model names (native SDK uses bare names)
- Preserves service account key auth via credentials parameter
- Delete vertexai_token_refresher.py (no longer needed)
- Strip markdown code fences in consolidator JSON parsing
- Rewrite vertexai tests for native SDK integration
2026-01-30 08:35:59 +01:00
..
consolidation Switch Vertex AI provider to native genai SDK (#242) 2026-01-30 08:35:59 +01:00
directives feat: revisit mental models, directives and reflections (#179) 2026-01-22 17:13:16 +01:00
mental_models feat: revisit mental models, directives and reflections (#179) 2026-01-22 17:13:16 +01:00
reflect fix: search_mental_models uuid type mismatch after text id migration (#225) 2026-01-28 20:04:19 +01:00
retain feat: add more config options for llm retries (#234) 2026-01-29 17:50:43 +01:00
search fix: graph endpoint not showing links for observations (#214) 2026-01-28 14:51:25 +01:00
__init__.py feat: extensions (#54) 2025-12-22 11:05:23 +01:00
cross_encoder.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
db_budget.py fix: batch queries on recall (#149) 2026-01-13 13:20:22 +01:00
db_utils.py misc: performance improvements (#140) 2026-01-09 14:47:20 +01:00
embeddings.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
entity_resolver.py fix: misc perf improvements (#133) 2026-01-08 22:49:04 +01:00
interface.py feat: introduce mental models (#132) 2026-01-16 11:16:41 +01:00
llm_wrapper.py Switch Vertex AI provider to native genai SDK (#242) 2026-01-30 08:35:59 +01:00
memory_engine.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
query_analyzer.py fix: improve mpfp retrieval (#146) 2026-01-12 18:58:05 +01:00
response_models.py fix: misc fixes for observations and mental models (#209) 2026-01-27 15:37:57 +01:00
task_backend.py fix: multi-tenant schema context for worker task execution (#208) 2026-01-27 12:28:47 +01:00
utils.py chore: drop dead code (#210) 2026-01-27 15:03:25 +01:00