SharePoint
Crawl sites, lists and pages through Microsoft Graph — extracting Word, Excel and PDF content — or sync document folders.
Turn your web estate, SharePoint, cloud drives and documents into knowledge agents can use — kept current on a schedule, searched with hybrid retrieval and cited with relevance scores on every answer.
Point Gecko at a site and it crawls, extracts and embeds the content — including modern JavaScript-rendered pages and the PDFs they link to — then keeps it current on a schedule.
^https://www\.university\.ac\.uk/(study|students)/**/news/****/staff-intranet/**nav.site-footerPolicies in SharePoint, handbooks in Google Drive, procedures in OneDrive — sync the folders that matter and keep them current automatically.
Crawl sites, lists and pages through Microsoft Graph — extracting Word, Excel and PDF content — or sync document folders.
Sync folders including Google Docs, Sheets and Slides alongside PDFs, Word and text files.
Bring in document folders from the storage your teams already use — including S3-compatible storage.
Choose a folder, include sub-folders if you want them, and set a sync schedule. Each source shows whether it is idle, pending, syncing, succeeded or failed.
For content that needs a human author, the knowledge base gives teams folders and a rich editor — and accepts documents straight from the desktop.
Deposits are refunded within 28 days of check-out, less any agreed deductions.
| Hall | Weekly rent |
|---|---|
| Riverside | £168 |
| Parkview | £142 |
Semantic search alone misses course codes, policy numbers and names. Gecko blends meaning with exact matching, then gives you the tools to tune precision and recall.
Rewrite the question into broader search variants so relevant content isn’t missed.
Expand User QueryHybrid by default — vector similarity fused with full-text search using reciprocal rank fusion.
Knowledge · WebsitesRe-score results with a dedicated reranking model from Cohere, Amazon Bedrock or a compatible endpoint.
Rerank ResultsMerge what matters and keep it inside a context budget before it reaches the model.
Consolidate ContextPer-agent minimum scores discard weak results before the model ever sees them.
Optionally split website content by meaning rather than fixed length, per site.
Full-text search catches the course code, the form number and the building name that embeddings blur.
Every crawled page and knowledge item records how often agents use it. Every answer records the sources behind it. Content ROI stops being a guess.
Point it at your website on Monday; answers cite live pages with relevance scores by Tuesday — and you can see exactly which content nobody needed.
Agents and workflow steps are granted specific knowledge folders and websites. The finance agent reads finance content; the admissions agent never touches HR policy.
Choose exactly which knowledge folders each agent or retrieval step can search.
Select the crawled websites an agent may search, with its own result limit and relevance floor for each.
Custom roles can be limited to specific websites, so a department manages its own content and nothing else.
Bring a website or a SharePoint site. We’ll show you how it would ground an agent, and which content would do the heavy lifting.