fleet-memory/hindsight-api/hindsight_api
Chris Bartholomew 49ae55af03
Switch Vertex AI provider to native genai SDK (#242)
Replace the OpenAI-compatible endpoint approach with the native
google-genai SDK for Vertex AI. This eliminates the custom token
refresher, TokenInjectingTransport, and async lifecycle complexity
while also removing the 8192 output token cap that the OpenAI
endpoint enforced.

Changes:
- vertexai provider now uses genai.Client(vertexai=True) instead of
  AsyncOpenAI with token-injecting transport
- Routes through existing _call_gemini/_call_with_tools_gemini paths
- Strips google/ prefix from model names (native SDK uses bare names)
- Preserves service account key auth via credentials parameter
- Delete vertexai_token_refresher.py (no longer needed)
- Strip markdown code fences in consolidator JSON parsing
- Rewrite vertexai tests for native SDK integration
2026-01-30 08:35:59 +01:00
..
admin feat: new 'worker' service (#176) 2026-01-20 10:17:56 +01:00
alembic fix: misc fixes for observations and mental models (#209) 2026-01-27 15:37:57 +01:00
api fix: /version endpoint return wrong version (#224) 2026-01-29 08:40:01 +01:00
engine Switch Vertex AI provider to native genai SDK (#242) 2026-01-30 08:35:59 +01:00
extensions feat: support different default pg schema (#222) 2026-01-28 18:14:44 +01:00
worker fix: run migrations on tenant schemas at startup and harden worker poller (#237) 2026-01-29 15:03:20 -05:00
__init__.py Release v0.4.2 2026-01-29 17:54:19 +01:00
banner.py fix: doc build and lint files (#34) 2025-12-16 13:49:09 +01:00
config.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
daemon.py fix: hindsight-embed on macos crashes (#228) 2026-01-29 16:57:51 +01:00
main.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
mcp_local.py feat(mcp): add timestamp to retain (#190) 2026-01-23 16:00:43 +01:00
mcp_tools.py feat(mcp): add timestamp to retain (#190) 2026-01-23 16:00:43 +01:00
metrics.py feat: add tenant to metrics labels (#151) 2026-01-13 15:31:58 +01:00
migrations.py fix: batch queries on recall (#149) 2026-01-13 13:20:22 +01:00
models.py chore: drop unused access_count column (#178) 2026-01-20 10:36:02 +01:00
pg0.py feat: support vertex as llm provider (#233) 2026-01-29 16:13:57 -05:00
server.py Fix: Load extensions in server.py for multi-worker deployments (#155) 2026-01-13 17:55:33 +01:00