Back to Support
onboarding

Uploading Documents

Learn how to upload documents, paste text, import from the web, or import GitHub issues.

Upstream turns your documents into structured intelligence. There are four ways to add content to your knowledge base, and each one feeds the same processing pipeline.

Four ways to add content

1. File Drop

Drag and drop files directly onto the upload area, or click Browse to select files from your computer. Supported file types:

  • PDF — Sales decks, research reports, strategy documents
  • DOCX — Word documents, meeting notes, proposals
  • TXT — Plain text exports, call transcripts, raw notes

You can upload multiple files at once. Each file is processed independently.

2. Paste Text

Click the Paste Text tab in the upload flow and paste content directly into the text box. This is ideal for:

  • Call notes you typed during a meeting
  • Email threads with customer feedback
  • Slack messages or chat logs
  • Quick notes that do not exist as a file

Give your pasted content a descriptive title so you can find it later.

3. Web Import

Click the Web Import tab, enter a URL, and Upstream will extract the content from the page. Use this for:

  • Blog posts or news articles about competitors
  • Public company pages or product announcements
  • Industry reports published online
  • Any publicly accessible web page with relevant content

4. GitHub Issues

In the GitHub Issues section, enter a repository as owner/name, choose open or all issues, optionally set an "updated since" date, and click Import issues. For private repositories, paste a GitHub token with read access to issues. The token is used for that one request and is never stored.

Each issue (title, description, labels, and comments) becomes a document, so extracted entities such as feature requests and pain points link straight back to the issue. Importing again only picks up new or edited issues. Once a set of issues is indexed, an Product Insights artifact can be generated that groups the extracted entities into tagged, weighted themes; large sets are split into an overview plus one artifact per theme.

What happens after you upload

Once you submit content through any of the four methods, Upstream processes it through several stages:

  1. Parsing — The document is read and converted into clean text.
  2. Entity extraction — Upstream identifies people, companies, products, themes, and other entities mentioned in the content.
  3. Relationship mapping — Connections between entities are detected (e.g., "Jane works at Acme" creates a relationship between the person and the company).
  4. Embedding generation — The content is indexed so it can be searched and retrieved when you ask questions in Chat.
  5. Review — Newly extracted entities appear in your review queue so you can confirm, edit, or dismiss them.

Tracking upload status

Visit the Upload Status page to see where each document is in the pipeline. Each upload shows its current stage, and you will see a confirmation when processing is complete. If something goes wrong during processing, the status page will show an error with details.

Tips for best results

  • Use descriptive file names. "Q4 Competitive Analysis - Acme Corp.pdf" is far more useful than "doc1.pdf" when you search later.
  • Upload regularly. The more content Upstream has, the richer your knowledge graph becomes and the better your chat answers get.
  • Mix your sources. Call transcripts, strategy docs, customer feedback, and competitive research together give you the most complete picture.
  • Keep files focused. A single-topic document produces cleaner entity extraction than a 50-page catch-all dump.
  • Review extracted entities promptly. Confirming or dismissing entities after upload keeps your knowledge base accurate and improves future extractions.

Still need help?

Email us at hello@getupstream.ai — we read every message.

Still need help? Contact our support team