Açıklama Yok

Deshraj Yadav cc1ee1deaa Fix dependencies in base package (#863) 1 yıl önce
.github 2b881aaad0 Cache dependencies in CI (#832) 1 yıl önce
configs 0f8a2e624a [Improvement] Add support for gpt4all through langchain (#838) 1 yıl önce
docs f6c4f86986 [Pipelines] Improvements in pipelines feature (#861) 1 yıl önce
embedchain 29bd038579 [Bug fix] Fix missing dependency issue with gmail (#862) 1 yıl önce
embedchain-js 3c3d98b9c3 feat: add embedchain javascript package (#576) 1 yıl önce
examples 78ec91a3a9 [Bug fix] Fix sqlite related issue with api server example (#857) 1 yıl önce
notebooks b2286f3e34 Google Colab Notebooks for LLMs, Embedders and VectorDBs (#821) 1 yıl önce
tests 29bd038579 [Bug fix] Fix missing dependency issue with gmail (#862) 1 yıl önce
.env.example 702067e521 Improve user readability of .env.example (#781) 1 yıl önce
.gitignore d77e8da3f3 [Feature] Update `db.query` to return source of context (#831) 1 yıl önce
.pre-commit-config.yaml ac68986404 Add project tools and contributing guidelines (#281) 2 yıl önce
CITATION.cff 736b645fea Add citation (#137) 2 yıl önce
CONTRIBUTING.md 76f1993e7a Update CONTRIBUTING.md (#845) 1 yıl önce
LICENSE 65d1ff37e8 Create LICENSE 2 yıl önce
Makefile 7641cba01d [Feature] JSON data loader support (#816) 1 yıl önce
README.md 191ae3ec1e Update README (#855) 1 yıl önce
poetry.lock cc1ee1deaa Fix dependencies in base package (#863) 1 yıl önce
poetry.toml ac68986404 Add project tools and contributing guidelines (#281) 2 yıl önce
pyproject.toml cc1ee1deaa Fix dependencies in base package (#863) 1 yıl önce

README.md

embedchain

PyPI Slack Discord Twitter Substack Open in Colab codecov

Embedchain is a Data Platform for LLMs - load, index, retrieve, and sync any unstructured data. Using embedchain, you can easily create LLM powered apps over any data. If you want a javascript version, check out embedchain-js

Community

  • Join embedchain community on slack by accepting this invite

🤝 Schedule a 1-on-1 Session

Book a 1-on-1 Session with Taranjeet, the founder, to discuss any issues, provide feedback, or explore how we can improve Embedchain for you.

🔧 Quick install

pip install --upgrade embedchain

🔍 Demo

Try out embedchain in your browser:

Open in Colab

📖 Documentation

The documentation for embedchain can be found at docs.embedchain.ai.

💻 Usage

Embedchain empowers you to create ChatGPT like apps, on your own dynamic dataset.

Data types supported

  • Youtube video
  • PDF file
  • CSV file
  • Web page
  • MDX file
  • XML file
  • Sitemap
  • Doc file
  • Notion
  • JSON file
  • OpenAPI specs
  • Code docs website
  • Unstructured file loader and many more

You can find the full list of data types on our documentation.

Queries

For example, you can use Embedchain to create an Elon Musk bot using the following code:

import os
from embedchain import App

# Create a bot instance
os.environ["OPENAI_API_KEY"] = "YOUR API KEY"
elon_bot = App()

# Embed online resources
elon_bot.add("https://en.wikipedia.org/wiki/Elon_Musk")
elon_bot.add("https://www.forbes.com/profile/elon-musk")
elon_bot.add("https://www.youtube.com/watch?v=RcYjXbSJBN8")

# Query the bot
elon_bot.query("How many companies does Elon Musk run and name those?")
# Answer: Elon Musk currently runs several companies. As of my knowledge, he is the CEO and lead designer of SpaceX, the CEO and product architect of Tesla, Inc., the CEO and founder of Neuralink, and the CEO and founder of The Boring Company. However, please note that this information may change over time, so it's always good to verify the latest updates.

Examples

LLM Google Colab Replit
OpenAI Open In Colab Try with Replit Badge
Anthropic Open In Colab Try with Replit Badge
Azure OpenAI Open In Colab Try with Replit Badge
VertexAI Open In Colab Try with Replit Badge
Cohere Open In Colab Try with Replit Badge
Hugging Face Open In Colab Try with Replit Badge
JinaChat Open In Colab Try with Replit Badge
GPT4All Open In Colab Try with Replit Badge
Llama2 Open In Colab Try with Replit Badge
Embedding model Google Colab Replit
OpenAI Open In Colab Try with Replit Badge
VertexAI Open In Colab Try with Replit Badge
GPT4All Open In Colab Try with Replit Badge
Hugging Face Open In Colab Try with Replit Badge
Vector DB Google Colab Replit
ChromaDB Open In Colab Try with Replit Badge
Elasticsearch Open In Colab Try with Replit Badge
Opensearch Open In Colab Try with Replit Badge
Pinecone Open In Colab Try with Replit Badge

🤝 Contributing

Contributions are welcome! Please check out the issues on the repository, and feel free to open a pull request. For more information, please see the contributing guidelines.

For more reference, please go through Development Guide and Documentation Guide.

Telemetry

We collect anonymous usage metrics to enhance our package's quality and user experience. This includes data like feature usage frequency and system info, but never personal details. The data helps us prioritize improvements and ensure compatibility. If you wish to opt-out, set the app.config.collect_metrics = False in the code. We prioritize data security and don't share this data externally.

Citation

If you utilize this repository, please consider citing it with:

@misc{embedchain,
  author = {Taranjeet Singh, Deshraj Yadav},
  title = {Embedchain: Data platform for LLMs - load, index, retrieve, and sync any unstructured data},
  year = {2023},
  publisher = {GitHub},
  journal = {GitHub repository},
  howpublished = {\url{https://github.com/embedchain/embedchain}},
}