It would help if there was a node for: Lemonade Server
My use case:
Run LLMs on device with GPU and NPU acceleration across various backends such as onnxruntime, llamacpp (vulkan, rocm) and others in the context of n8n workflows
Any resources to support this?
GitHub repo: GitHub - lemonade-sdk/lemonade: Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk ยท GitHub
Website: lemonade-server.ai
Are you willing to work on this?
Yes, happy to make the contribution