langchain

mirror of https://github.com/hwchase17/langchain synced 2024-11-04 06:00:26 +00:00

History

Armin Stepanyan 641efcf41c community: add runtime kwargs to HuggingFacePipeline (#17005 ) This PR enables changing the behaviour of huggingface pipeline between different calls. For example, before this PR there's no way of changing maximum generation length between different invocations of the chain. This is desirable in cases, such as when we want to scale the maximum output size depending on a dynamic prompt size. Usage example: ```python from langchain_community.llms.huggingface_pipeline import HuggingFacePipeline from transformers import AutoModelForCausalLM, AutoTokenizer, pipeline model_id = "gpt2" tokenizer = AutoTokenizer.from_pretrained(model_id) model = AutoModelForCausalLM.from_pretrained(model_id) pipe = pipeline("text-generation", model=model, tokenizer=tokenizer) hf = HuggingFacePipeline(pipeline=pipe) hf("Say foo:", pipeline_kwargs={"max_new_tokens": 42}) ``` --------- Co-authored-by: Bagatur <baskaryan@gmail.com>		2024-02-08 13:58:31 -08:00
..
examples	community[minor]: New documents loader for visio files (with extension .vsdx) (#16171 )	2024-01-22 22:07:03 -08:00
integration_tests	community: add runtime kwargs to HuggingFacePipeline (#17005 )	2024-02-08 13:58:31 -08:00
unit_tests	community: Add you.com utility, update you retriever integration docs (#17014 )	2024-02-08 13:47:50 -08:00
__init__.py	community[major], core[patch], langchain[patch], experimental[patch]: Create langchain-community (#14463 )	2023-12-11 13:53:30 -08:00