Release· Databricks release notes (AWS)

Lakebase Search gets lakebase_tokenizer for configurable keyword matching

A new lakebase_tokenizer extension controls how Lakebase Search splits text into terms: normalization and accent stripping, your own stopword and synonym lists, and stemming. It produces standard tsvector values, so it works with to_tsvector, GIN indexes and the lakebase_bm25 index; install it after enabling Lakebase Search in project settings.

Read the original

What to re-read