Skip to content

Pull requests: NVIDIA/FasterTransformer

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobodyLoading
Sort

Pull requests list

Include stdio.h
#770 opened Oct 19, 2023 by JihaoXinLoading…
Ft llama opt
#762 opened Oct 2, 2023 by dypshongLoading…
Support Seq length up to 8K
#756 opened Sep 4, 2023 by zhen-jiaLoading…
[Bugfix] GptJ & GptNeoX batch inference error
#742 opened Aug 11, 2023 by YZP17121579Loading…
Add fusion-for-decoder-only for llama
#733 opened Jul 28, 2023 by binxuanLoading…
Fix beam search output_log_prob index error
#732 opened Jul 25, 2023 by cpm0722Loading…
Add cuDNN include path as a common include dir
#724 opened Jul 18, 2023 by jacobkahnLoading…
Remove parenthesis from asserts
#699 opened Jul 2, 2023 by miguelusqueLoading…
[Doc] Fix typo in gpt_guide.md
#682 opened Jun 26, 2023 by myry96Loading…
gptneox & gptj int8 quantization & share context
#653 opened Jun 7, 2023 by rahuanLoading…
Add missing headers
#648 opened Jun 1, 2023 by brian14708Loading…
Fix TOC of gptneox_guide.md
#633 opened May 23, 2023 by xu-songLoading…
fix multi-gpu build
#616 opened May 17, 2023 by dskhudiaContributorLoading…
Fix mpi library linking issue
#612 opened May 16, 2023 by liangfuLoading…
remove buggy duplicated code
#609 opened May 12, 2023 by chenho74Loading…
ProTip! Exclude everything labeled bug with -label:bug.