Skip to content

FEAT: Support TensorRT-LLM backend#646

Draft
aresnow1 wants to merge 4 commits intoxorbitsai:mainfrom
aresnow1:feat/tensorrt-llm
Draft

FEAT: Support TensorRT-LLM backend#646
aresnow1 wants to merge 4 commits intoxorbitsai:mainfrom
aresnow1:feat/tensorrt-llm

Conversation

@aresnow1
Copy link
Copy Markdown
Contributor

Support TensorRT-LLM backend.

  • Implements TRTModel with generate method.
  • Expose launch_trt_model to client.
  • Doc and example

@XprobeBot XprobeBot added this to the v0.6.3 milestone Nov 14, 2023
@aresnow1
Copy link
Copy Markdown
Contributor Author

Python API of in-flight batching is needed for this PR, and TensorRT-LLM team says it will be implemented in next versions.

@XprobeBot XprobeBot modified the milestones: v0.9.3, v0.9.4, v0.9.5 Mar 15, 2024
@XprobeBot XprobeBot modified the milestones: v0.10.0, v0.10.1 Mar 29, 2024
@XprobeBot XprobeBot modified the milestones: v0.10.1, v0.10.2 Apr 12, 2024
@XprobeBot XprobeBot modified the milestones: v0.10.2, v0.10.3, v0.11.0 Apr 19, 2024
@XprobeBot XprobeBot modified the milestones: v0.11.0, v0.11.1, v0.11.2 May 11, 2024
@XprobeBot XprobeBot modified the milestones: v0.11.2, v0.11.3 May 24, 2024
@XprobeBot XprobeBot modified the milestones: v0.11.3, v0.11.4, v0.12.0, v0.12.1 May 31, 2024
@XprobeBot XprobeBot modified the milestones: v0.12.1, v0.12.2 Jun 14, 2024
@XprobeBot XprobeBot modified the milestones: v0.12.2, v0.12.4, v0.13.0, v0.13.1 Jun 28, 2024
@XprobeBot XprobeBot modified the milestones: v0.13.1, v0.13.2 Jul 12, 2024
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants