Skip to content

feat(extra-backends): Improvements, adding mamba example - #1618

Merged
mudler merged 3 commits into
masterfrom
mamba_example
Jan 20, 2024
Merged

feat(extra-backends): Improvements, adding mamba example#1618
mudler merged 3 commits into
masterfrom
mamba_example

Conversation

@mudler

@mudler mudler commented Jan 20, 2024

Copy link
Copy Markdown
Owner

vllm: add max_tokens, wire up stream event (fixes #1609)
mamba: fixups, adding examples for mamba-chat

Description

This PR is part of #1588

Notes for Reviewers

Signed commits

  • Yes, I signed my commits.

vllm: add max_tokens, wire up stream event
mamba: fixups, adding examples for mamba-chat
@netlify

netlify Bot commented Jan 20, 2024

Copy link
Copy Markdown

Deploy Preview for localai canceled.

Name Link
🔨 Latest commit 1fdb857
🔍 Latest deploy log https://app.netlify.com/sites/localai/deploys/65abe30da33a1800082e9958

@netlify

netlify Bot commented Jan 20, 2024

Copy link
Copy Markdown

Deploy Preview for localai ready!

Name Link
🔨 Latest commit 45c86b1
🔍 Latest deploy log https://app.netlify.com/sites/localai/deploys/65abe6dcc66be20008feb242
😎 Deploy Preview https://deploy-preview-1618--localai.netlify.app
📱 Preview on mobile
Toggle QR Code...

QR Code

Use your smartphone camera to open QR code link.

To edit notification comments on pull requests, go to your Netlify site configuration.

@mudler mudler mentioned this pull request Jan 20, 2024
1 task
@mudler
mudler merged commit 06cd9ef into master Jan 20, 2024
@mudler
mudler deleted the mamba_example branch January 20, 2024 16:56
@mudler mudler added the enhancement New feature or request label Jan 20, 2024
truecharts-admin referenced this pull request in trueforge-org/truecharts Jan 21, 2024
….0 by renovate (#17442)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [docker.io/localai/localai](https://togithub.com/mudler/LocalAI) |
minor | `v2.5.1-cublas-cuda11-ffmpeg-core` ->
`v2.6.0-cublas-cuda11-ffmpeg-core` |

---

> [!WARNING]
> Some dependencies could not be looked up. Check the Dependency
Dashboard for more information.

---

### Release Notes

<details>
<summary>mudler/LocalAI (docker.io/localai/localai)</summary>

### [`v2.6.0`](https://togithub.com/mudler/LocalAI/releases/tag/v2.6.0)

[Compare
Source](https://togithub.com/mudler/LocalAI/compare/v2.5.1...v2.6.0)

<!-- Release notes generated using configuration in .github/release.yml
at master -->

##### What's Changed

##### Bug fixes 🐛

- move BUILD_GRPC_FOR_BACKEND_LLAMA logic to makefile: errors in this
section now immediately fail the build by
[@&#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1576](https://togithub.com/mudler/LocalAI/pull/1576)
- prepend built binaries in PATH for BUILD_GRPC_FOR_BACKEND_LLAMA by
[@&#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1593](https://togithub.com/mudler/LocalAI/pull/1593)

##### Exciting New Features 🎉

- minor: replace shell pwd in Makefile with CURDIR for better windows
compatibility by [@&#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1571](https://togithub.com/mudler/LocalAI/pull/1571)
- Makefile: allow to build without GRPC_BACKENDS by
[@&#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1607](https://togithub.com/mudler/LocalAI/pull/1607)
- feat: 🐍 add mamba support by
[@&#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1589](https://togithub.com/mudler/LocalAI/pull/1589)
- feat(extra-backends): Improvements, adding mamba example by
[@&#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1618](https://togithub.com/mudler/LocalAI/pull/1618)

##### 👒 Dependencies

- ⬆️ Update docs version mudler/LocalAI by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1567](https://togithub.com/mudler/LocalAI/pull/1567)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1568](https://togithub.com/mudler/LocalAI/pull/1568)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1573](https://togithub.com/mudler/LocalAI/pull/1573)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1578](https://togithub.com/mudler/LocalAI/pull/1578)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1583](https://togithub.com/mudler/LocalAI/pull/1583)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1587](https://togithub.com/mudler/LocalAI/pull/1587)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1590](https://togithub.com/mudler/LocalAI/pull/1590)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1594](https://togithub.com/mudler/LocalAI/pull/1594)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1599](https://togithub.com/mudler/LocalAI/pull/1599)

##### Other Changes

- Moving the how tos to self hosted by
[@&#8203;lunamidori5](https://togithub.com/lunamidori5) in
[https://github.com/mudler/LocalAI/pull/1574](https://togithub.com/mudler/LocalAI/pull/1574)
- docs: missing golang requirement for local build for debian by
[@&#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1596](https://togithub.com/mudler/LocalAI/pull/1596)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1597](https://togithub.com/mudler/LocalAI/pull/1597)
- docs/examples: enhancements by
[@&#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1572](https://togithub.com/mudler/LocalAI/pull/1572)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1604](https://togithub.com/mudler/LocalAI/pull/1604)
- Update README.md by
[@&#8203;lunamidori5](https://togithub.com/lunamidori5) in
[https://github.com/mudler/LocalAI/pull/1601](https://togithub.com/mudler/LocalAI/pull/1601)
- docs: re-use original permalinks by
[@&#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1610](https://togithub.com/mudler/LocalAI/pull/1610)
- ⬆️ Update ggerganov/llama.cpp by
[@&#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1612](https://togithub.com/mudler/LocalAI/pull/1612)
- Expanded and interlinked Docker documentation by
[@&#8203;jamesbraza](https://togithub.com/jamesbraza) in
[https://github.com/mudler/LocalAI/pull/1614](https://togithub.com/mudler/LocalAI/pull/1614)
- Modernized LlamaIndex integration by
[@&#8203;jamesbraza](https://togithub.com/jamesbraza) in
[https://github.com/mudler/LocalAI/pull/1613](https://togithub.com/mudler/LocalAI/pull/1613)

##### New Contributors

- [@&#8203;dionysius](https://togithub.com/dionysius) made their first
contribution in
[https://github.com/mudler/LocalAI/pull/1571](https://togithub.com/mudler/LocalAI/pull/1571)

**Full Changelog**:
mudler/LocalAI@v2.5.1...v2.6.0

</details>

---

### Configuration

📅 **Schedule**: Branch creation - "before 10pm on monday" in timezone
Europe/Amsterdam, Automerge - At any time (no schedule defined).

🚦 **Automerge**: Enabled.

♻ **Rebasing**: Whenever PR becomes conflicted, or you tick the
rebase/retry checkbox.

🔕 **Ignore**: Close this PR and you won't be reminded about this update
again.

---

- [ ] <!-- rebase-check -->If you want to rebase/retry this PR, check
this box

---

This PR has been generated by [Renovate
Bot](https://togithub.com/renovatebot/renovate).

<!--renovate-debug:eyJjcmVhdGVkSW5WZXIiOiIzNy4xNDAuMTYiLCJ1cGRhdGVkSW5WZXIiOiIzNy4xNDAuMTYiLCJ0YXJnZXRCcmFuY2giOiJtYXN0ZXIifQ==-->
GabrielBarzen referenced this pull request in GabrielBarzen/charts Feb 2, 2024
….0 by renovate (trueforge-org#17442)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [docker.io/localai/localai](https://togithub.com/mudler/LocalAI) |
minor | `v2.5.1-cublas-cuda11-ffmpeg-core` ->
`v2.6.0-cublas-cuda11-ffmpeg-core` |

---

> [!WARNING]
> Some dependencies could not be looked up. Check the Dependency
Dashboard for more information.

---

### Release Notes

<details>
<summary>mudler/LocalAI (docker.io/localai/localai)</summary>

### [`v2.6.0`](https://togithub.com/mudler/LocalAI/releases/tag/v2.6.0)

[Compare
Source](https://togithub.com/mudler/LocalAI/compare/v2.5.1...v2.6.0)

<!-- Release notes generated using configuration in .github/release.yml
at master -->

##### What's Changed

##### Bug fixes 🐛

- move BUILD_GRPC_FOR_BACKEND_LLAMA logic to makefile: errors in this
section now immediately fail the build by
[@&trueforge-org#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1576](https://togithub.com/mudler/LocalAI/pull/1576)
- prepend built binaries in PATH for BUILD_GRPC_FOR_BACKEND_LLAMA by
[@&trueforge-org#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1593](https://togithub.com/mudler/LocalAI/pull/1593)

##### Exciting New Features 🎉

- minor: replace shell pwd in Makefile with CURDIR for better windows
compatibility by [@&trueforge-org#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1571](https://togithub.com/mudler/LocalAI/pull/1571)
- Makefile: allow to build without GRPC_BACKENDS by
[@&trueforge-org#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1607](https://togithub.com/mudler/LocalAI/pull/1607)
- feat: 🐍 add mamba support by
[@&trueforge-org#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1589](https://togithub.com/mudler/LocalAI/pull/1589)
- feat(extra-backends): Improvements, adding mamba example by
[@&trueforge-org#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1618](https://togithub.com/mudler/LocalAI/pull/1618)

##### 👒 Dependencies

- ⬆️ Update docs version mudler/LocalAI by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1567](https://togithub.com/mudler/LocalAI/pull/1567)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1568](https://togithub.com/mudler/LocalAI/pull/1568)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1573](https://togithub.com/mudler/LocalAI/pull/1573)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1578](https://togithub.com/mudler/LocalAI/pull/1578)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1583](https://togithub.com/mudler/LocalAI/pull/1583)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1587](https://togithub.com/mudler/LocalAI/pull/1587)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1590](https://togithub.com/mudler/LocalAI/pull/1590)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1594](https://togithub.com/mudler/LocalAI/pull/1594)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1599](https://togithub.com/mudler/LocalAI/pull/1599)

##### Other Changes

- Moving the how tos to self hosted by
[@&trueforge-org#8203;lunamidori5](https://togithub.com/lunamidori5) in
[https://github.com/mudler/LocalAI/pull/1574](https://togithub.com/mudler/LocalAI/pull/1574)
- docs: missing golang requirement for local build for debian by
[@&trueforge-org#8203;dionysius](https://togithub.com/dionysius) in
[https://github.com/mudler/LocalAI/pull/1596](https://togithub.com/mudler/LocalAI/pull/1596)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1597](https://togithub.com/mudler/LocalAI/pull/1597)
- docs/examples: enhancements by
[@&trueforge-org#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1572](https://togithub.com/mudler/LocalAI/pull/1572)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1604](https://togithub.com/mudler/LocalAI/pull/1604)
- Update README.md by
[@&trueforge-org#8203;lunamidori5](https://togithub.com/lunamidori5) in
[https://github.com/mudler/LocalAI/pull/1601](https://togithub.com/mudler/LocalAI/pull/1601)
- docs: re-use original permalinks by
[@&trueforge-org#8203;mudler](https://togithub.com/mudler) in
[https://github.com/mudler/LocalAI/pull/1610](https://togithub.com/mudler/LocalAI/pull/1610)
- ⬆️ Update ggerganov/llama.cpp by
[@&trueforge-org#8203;localai-bot](https://togithub.com/localai-bot) in
[https://github.com/mudler/LocalAI/pull/1612](https://togithub.com/mudler/LocalAI/pull/1612)
- Expanded and interlinked Docker documentation by
[@&trueforge-org#8203;jamesbraza](https://togithub.com/jamesbraza) in
[https://github.com/mudler/LocalAI/pull/1614](https://togithub.com/mudler/LocalAI/pull/1614)
- Modernized LlamaIndex integration by
[@&trueforge-org#8203;jamesbraza](https://togithub.com/jamesbraza) in
[https://github.com/mudler/LocalAI/pull/1613](https://togithub.com/mudler/LocalAI/pull/1613)

##### New Contributors

- [@&trueforge-org#8203;dionysius](https://togithub.com/dionysius) made their first
contribution in
[https://github.com/mudler/LocalAI/pull/1571](https://togithub.com/mudler/LocalAI/pull/1571)

**Full Changelog**:
mudler/LocalAI@v2.5.1...v2.6.0

</details>

---

### Configuration

📅 **Schedule**: Branch creation - "before 10pm on monday" in timezone
Europe/Amsterdam, Automerge - At any time (no schedule defined).

🚦 **Automerge**: Enabled.

♻ **Rebasing**: Whenever PR becomes conflicted, or you tick the
rebase/retry checkbox.

🔕 **Ignore**: Close this PR and you won't be reminded about this update
again.

---

- [ ] <!-- rebase-check -->If you want to rebase/retry this PR, check
this box

---

This PR has been generated by [Renovate
Bot](https://togithub.com/renovatebot/renovate).

<!--renovate-debug:eyJjcmVhdGVkSW5WZXIiOiIzNy4xNDAuMTYiLCJ1cGRhdGVkSW5WZXIiOiIzNy4xNDAuMTYiLCJ0YXJnZXRCcmFuY2giOiJtYXN0ZXIifQ==-->
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

enhancement New feature or request

Projects

None yet

Development

Successfully merging this pull request may close these issues.

When using request to vllm backend with stream = true, no actual result is returned

1 participant