timothycarambat
c65ab6d863
merge with master
2024-05-21 14:48:16 -05:00
Timothy Carambat
7e0b638a2c
Patch confluence URL patterns( #1426 )
...
* patch confluence patterns
---------
Co-authored-by: shatfield4 <seanhatfield5@gmail.com>
2024-05-16 14:15:59 -07:00
timothycarambat
87b41a60e9
refactor spaceKey url pattern for custom domains
2024-05-16 11:01:34 -07:00
Predrag Stojadinović
cf969adf37
1362 custom display confluence url ( #1423 )
...
* chore: confluence data connector can now handle custom urls, in addition to default {subdomain}.atlassian.net ones
* chore: formatting as per yarn lint
* chore: adding /display/ url matching to confluence data connector
2024-05-16 10:46:18 -07:00
timothycarambat
d603d0fd51
patch:update storage for bulk-website scraper for render
2024-05-14 12:59:14 -07:00
timothycarambat
c8dac6177a
Merge branch 'master' of github.com:Mintplex-Labs/anything-llm into render
2024-05-14 12:57:44 -07:00
timothycarambat
b5ac944475
patch: bulk-scraper, update when folder is made and path creation params
2024-05-14 12:57:23 -07:00
timothycarambat
72c9fda6c9
Merge branch 'master' of github.com:Mintplex-Labs/anything-llm into render
2024-05-14 12:50:17 -07:00
Sean Hatfield
612a7e1662
[FEAT] Website depth scraping data connector ( #1191 )
...
* WIP website depth scraping, (sort of works)
* website depth data connector stable + add maxLinks option
* linting + loading small ui tweak
* refactor website depth data connector for stability, speed, & readability
* patch: remove console log
Guard clause on URL validitiy check
reasonable overrides
---------
Co-authored-by: Timothy Carambat <rambat1010@gmail.com>
2024-05-14 12:49:14 -07:00
jazelly
d71db22799
fix: skip undefined confluence pageContent ( #1383 )
...
Refs: https://github.com/Mintplex-Labs/anything-llm/issues/1381
Co-authored-by: Timothy Carambat <rambat1010@gmail.com>
2024-05-14 10:22:13 -07:00
Predrag Stojadinović
78e3e35d27
[FEAT] Confluence Data Connector handles custom Confluence urls ( #1362 )
...
* chore: confluence data connector can now handle custom urls, in addition to default {subdomain}.atlassian.net ones
* chore: formatting as per yarn lint
2024-05-14 10:21:04 -07:00
timothycarambat
c60077a078
merge with master
2024-05-03 10:02:53 -07:00
timothycarambat
2d215acb75
patch storage dirs for extensions
2024-05-02 14:03:10 -07:00
timothycarambat
6150ff41ea
Merge branch 'master' of github.com:Mintplex-Labs/anything-llm into render
2024-05-01 13:33:07 -07:00
Sean Hatfield
348b36bf85
[FEAT] Confluence data connector ( #1181 )
...
* WIP Confluence data connector backend
* confluence data connector complete
* confluence citations
* fix citation for confluence
* Patch confulence integration
* fix Citation Icon for confluence
---------
Co-authored-by: timothycarambat <rambat1010@gmail.com>
2024-04-25 17:53:38 -07:00
timothycarambat
fde4e5400f
Merge branch 'master' of github.com:Mintplex-Labs/anything-llm into render
2024-04-12 14:57:46 -07:00
Sean Hatfield
af84b01482
[FIX] GitHub repo with periods in link fix ( #1084 )
...
fix periods in github repo links bug
2024-04-12 14:56:59 -07:00
timothycarambat
75ced7e65a
merge with master
...
Patch LLM selection for native to be disabled
2024-04-07 14:55:18 -07:00
Timothy Carambat
1f8ab0d245
Remove YoutubeLoader dependency ( #1050 )
...
* WIP data connector redesign
* new UI for data connectors complete
* remove old data connector page/cleanup imports
* cleanup of UI and imports
* Remove Youtube Transcript dep and move in-house
* lang pref default to en
---------
Co-authored-by: shatfield4 <seanhatfield5@gmail.com>
2024-04-05 16:33:01 -07:00
timothycarambat
ae01785220
Merge branch 'master' of github.com:Mintplex-Labs/anything-llm into render
2024-02-21 15:11:45 -08:00
Timothy Carambat
d89610586a
improve error messages from YT scraping ( #768 )
...
parse & enforce URL to allow multiple URL schemas
2024-02-21 10:47:10 -08:00
Timothy Carambat
d52f8aafd4
689 links in citation ( #715 )
...
* Include links in citations
force ChunkSource key to retain this information
old links will be unsupported
* show special icons depending on source
* remove console log
* reset server documents writeTo
2024-02-13 14:11:57 -08:00
Sean Hatfield
288ff0d18c
fix vector cache not deleting cache after unembedding items with folders ( #630 )
2024-01-22 13:03:05 -08:00
timothycarambat
0eb2fe7248
Map .env to storage .env file
...
map writeToServerDocuments to resolve to fixed storage mount for Render
2023-12-19 11:35:20 -08:00
Timothy Carambat
ecf4295537
Add ability to grab youtube transcripts via doc processor ( #470 )
...
* Add ability to grab youtube transcripts via doc processor
* dynamic imports
swap out Github for Youtube in placeholder text
2023-12-18 17:17:26 -08:00
Timothy Carambat
452582489e
GitHub loader extension + extension support v1 ( #469 )
...
* feat: implement github repo loading
fix: purge of folders
fix: rendering of sub-files
* noshow delete on custom-documents
* Add API key support because of rate limits
* WIP for frontend of data connectors
* wip
* Add frontend form for GitHub repo data connector
* remove console.logs
block custom-documents from being deleted
* remove _meta unused arg
* Add support for ignore pathing in request
Ignore path input via tagging
* Update hint
2023-12-18 15:48:02 -08:00