Content
Name
Ready to Publish
Tags
Publish Date
Slug
Featured
Authors
Excerpt
Extra Info
Last Edited Time
Related Posts
Do not index
Hide CTA
Hide in Main Feed
Meta Description
Meta Title
Hide Cover
Original Paper
Ready to Publish
Featured
Oct 15, 2024 07:28 PM
Ready to Publish
Featured
Aug 28, 2024 01:30 AM
Jun 15, 2024
Aug 23, 2024 02:53 AM
llmEngineer.weekly: Running models locally, LLM-graded evals too expensive for production? Here's our solution...RAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
Jun 12, 2024
Aug 23, 2024 02:47 AM
How the best teams evaluate their chatbotsCustom Evaluations for your AI for freeImproving Retrieval Augmented Language Model with Self-Reasoning
Jun 7, 2024
Aug 23, 2024 12:08 AM
Self-Taught EvaluatorsRAGEval: Scenario Specific RAG Evaluation Dataset Generation FrameworkAthina IDE: A Collaborative Editor for AI teams to Prototype, Evaluate, and Experiment
Jun 7, 2024
Aug 23, 2024 12:01 AM
Conversational Prompt EngineeringRAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented GenerationAthina IDE: A Collaborative Editor for AI teams to Prototype, Evaluate, and Experiment
Jul 1, 2024
Aug 22, 2024 11:58 PM
Are you afraid of making changes to your LLM pipeline?June Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + moreRAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
Jun 14, 2024
Aug 22, 2024 11:52 PM
Conversational Prompt EngineeringFrom LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and FutureJune Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + more
Jun 18, 2024
Aug 22, 2024 11:49 PM
Analyze and compare LLM performance across different prompts, models, and topicsCompare Mode on Athina IDECommon LLM chatbot problems and how to solve them
Jun 27, 2024
Aug 22, 2024 11:45 PM
From Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic DataSupport for Custom Models hosted on Azure and AWS Bedrock!From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
Jun 14, 2024
Aug 22, 2024 11:42 PM
Athina IDE: A Collaborative Editor for AI teams to Prototype, Evaluate, and ExperimentFrom LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and FutureJune Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + more
Jun 25, 2024
Aug 22, 2024 11:38 PM
June Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + moreAthina IDE: A Collaborative Editor for AI teams to Prototype, Evaluate, and ExperimentRAGEval: Scenario Specific RAG Evaluation Dataset Generation Framework
Jul 29, 2024
Aug 22, 2024 11:34 PM
Prompts, Prompts, Prompts!RAGEval: Scenario Specific RAG Evaluation Dataset Generation FrameworkSelf-Taught Evaluators
Jul 31, 2024
Aug 22, 2024 11:20 PM
How non-technical users can prototype pipelines, run AI experiments and evaluationsEvaluate llama-3 vs gpt-4o on YOUR dataset in a few clicksRe-run your production traces on different LLMs and compare the results
Aug 2, 2024
Aug 21, 2024 10:38 PM
Evaluating LLM Chatbot Conversations is hard - here's how we're solving itEvaluating JSON responses: LLMs still can't be trusted to produce consistent JSON outputsGenerate high-quality synthetic datasets for RAG Q&A in 30 secondsDiscovering Preference Optimization Algorithms with and for Large Language Models
Jul 29, 2024
Aug 21, 2024 10:28 PM
Have you tried Cohere's new model Command R+? Compare against Claude 3 and Gemini Pro on AthinaAthina IDE: A Collaborative Editor for AI teams to Prototype, Evaluate, and ExperimentJune Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + more
Aug 8, 2024
Aug 21, 2024 10:19 PM
Configuring an eval in 15 seconds (yes, really)June Product Updates: Enterprise Features, Dynamic Columns, Spreadsheet-ing, Prompt Management + moreFrom LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and FuturePersonaGym: Evaluating Persona Agents and LLMsSelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
Aug 18, 2024
Aug 21, 2024 10:15 PM
We just launched on Product Hunt!Annotate LLM traces on Athina + new models support, automatic token & cost trackingHow to backtest prompt / model changes?PersonaGym: Evaluating Persona Agents and LLMsFollowing Length Constraints in InstructionsTree Search For Language Model AgentsSelf-Tuning: Instructing LLMs to Effectively Acquire New Knowledge through Self-TeachingSelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
Aug 8, 2024
Aug 21, 2024 10:11 PM
From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and FutureConversational Prompt Engineering🛠️ Understand Your Users → Detect Hallucinations → IterateBe like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMsMixture-of-Agents Enhances Large Language Model Capabilities
Aug 5, 2024
Aug 21, 2024 10:07 PM
Product Hunt Launch: Help us get to #1 Product of the Day!Access your LLM Traces via our GraphQL APIRAG Foundry: A Framework for Enhancing LLMs for Retrieval Augmented GenerationMixture-of-Agents Enhances Large Language Model Capabilities
Aug 5, 2024
Aug 21, 2024 10:07 PM
LLM Critics Help Catch LLM BugsFrom LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and FutureCan I walk you through Athina in 15 mins?Conversational Prompt EngineeringSelf-Taught EvaluatorsOn LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A SurveyFrom Artificial Needles to Real Haystacks: Improving Retrieval Capabilities in LLMs by Finetuning on Synthetic DataBe like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMs
Ready to Publish
Featured
Aug 6, 2024 04:40 AM
Ready to Publish
Aug 3, 2024
Featured
Aug 3, 2024 10:10 PM
Jul 9, 2024
Jul 9, 2024 07:58 PM
Self-Taught EvaluatorsConcise Thoughts: Impact of Output Length on LLM Reasoning and CostFollowing Length Constraints in InstructionsOn LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A SurveyBe like a Goldfish, Don't Memorize! Mitigating Memorization in Generative LLMsTree Search For Language Model Agents
Nov 21, 2023
Jul 7, 2024 03:27 PM
Jan 30, 2024
Jul 7, 2024 03:10 PM
Jun 3, 2024
Jul 7, 2024 02:57 PM
Jul 7, 2024 02:31 PM
Jul 7, 2024 02:27 PM
Jan 24, 2024
Jul 7, 2024 02:25 PM
Jul 7, 2024 02:18 PM
Jul 7, 2024 02:16 PM
Mar 18, 2024
Jul 7, 2024 02:12 PM
Jul 7, 2024 02:05 PM
Jun 1, 2024
Jul 7, 2024 01:44 PM
Jun 5, 2024
Jul 7, 2024 01:43 PM
Concise Thoughts: Impact of Output Length on LLM Reasoning and CostFollowing Length Constraints in InstructionsOn LLMs-Driven Synthetic Data Generation, Curation, and Evaluation: A SurveyMixture-of-Agents Enhances Large Language Model CapabilitiesSelfGoal: Your Language Agents Already Know How to Achieve High-level Goals
May 9, 2024
Jul 7, 2024 01:42 PM
Jun 29, 2024
Jul 7, 2024 01:41 PM
May 25, 2024
Jul 7, 2024 01:39 PM
Apr 16, 2024
Jul 7, 2024 01:38 PM
Apr 30, 2024
Jul 7, 2024 01:37 PM
Jun 19, 2024
Jul 7, 2024 01:36 PM
May 16, 2024
Jul 7, 2024 01:36 PM
May 3, 2024
Jul 7, 2024 01:35 PM
Jun 10, 2024
Jul 7, 2024 01:35 PM
Jun 12, 2024
Jul 7, 2024 01:34 PM
Jun 26, 2024
Jul 7, 2024 01:34 PM
May 20, 2024
Jul 7, 2024 01:31 PM
Jun 2, 2024
Jul 7, 2024 01:27 PM
Apr 9, 2024
Jul 7, 2024 01:26 PM
Jan 17, 2024
Jul 7, 2024 01:26 PM
Mar 13, 2024
Jul 7, 2024 01:25 PM
Mar 18, 2024
Jul 7, 2024 01:24 PM
Mar 7, 2024
Jul 7, 2024 01:23 PM
Apr 24, 2024
Jul 5, 2024 03:22 PM
Ready to Publish
Dec 15, 2023
Jun 29, 2024 10:49 PM
Image-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningText-driven Prompt Generation for Vision-Language Models in Federated LearningReverse Stable Diffusion: What prompt was used to generate this image?
Ready to Publish
Mar 14, 2024
Jun 29, 2024 10:46 PM
FoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph PromptText-driven Prompt Generation for Vision-Language Models in Federated LearningConsistency-guided Prompt Learning for Vision-Language Models
Ready to Publish
Aug 8, 2023
Jun 29, 2024 10:41 PM
FoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph PromptText-driven Prompt Generation for Vision-Language Models in Federated LearningConsistency-guided Prompt Learning for Vision-Language Models
Ready to Publish
May 4, 2023
Jun 29, 2024 10:38 PM
Text-driven Prompt Generation for Vision-Language Models in Federated LearningImage-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningConsistency-guided Prompt Learning for Vision-Language Models
Ready to Publish
Apr 15, 2023
Jun 29, 2024 10:35 PM
Text-driven Prompt Generation for Vision-Language Models in Federated LearningImage-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningConsistency-guided Prompt Learning for Vision-Language Models
Ready to Publish
Aug 2, 2023
Jun 29, 2024 10:32 PM
Image-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningConsistency-guided Prompt Learning for Vision-Language ModelsReverse Stable Diffusion: What prompt was used to generate this image?Rethinking Visual Prompt Learning as Masked Visual Token Modeling
Ready to Publish
Feb 27, 2024
Jun 29, 2024 10:29 PM
LLM Critics Help Catch LLM BugsFoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph PromptImage-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningReverse Stable Diffusion: What prompt was used to generate this image?Does Prompt-Tuning Language Model Ensure Privacy?Prompt-ICM: A Unified Framework towards Image Coding for Machines with Task-driven PromptsLarge Language Model Prompt Chaining for Long Legal Document ClassificationPlum: Prompt Learning using Metaheuristic
Ready to Publish
Dec 7, 2023
Jun 29, 2024 10:26 PM
Image-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningText-driven Prompt Generation for Vision-Language Models in Federated LearningRobust Safety Classifier for Large Language Models: Adversarial Prompt ShieldConsistency-guided Prompt Learning for Vision-Language ModelsReverse Stable Diffusion: What prompt was used to generate this image?Does Prompt-Tuning Language Model Ensure Privacy?Prompt-ICM: A Unified Framework towards Image Coding for Machines with Task-driven PromptsRethinking Visual Prompt Learning as Masked Visual Token Modeling
Ready to Publish
Oct 9, 2023
Jun 29, 2024 10:21 PM
FoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph PromptText-driven Prompt Generation for Vision-Language Models in Federated Learningviz2viz: Prompt-driven stylized visualization generation using a diffusion modelImage-Object-Specific Prompt Learning for Few-Shot Class-Incremental LearningDoes Prompt-Tuning Language Model Ensure Privacy?Prompt-ICM: A Unified Framework towards Image Coding for Machines with Task-driven PromptsLarge Language Model Prompt Chaining for Long Legal Document ClassificationPlum: Prompt Learning using MetaheuristicRethinking Visual Prompt Learning as Masked Visual Token Modeling
Ready to Publish
Aug 20, 2023
Jun 29, 2024 10:18 PM
Prompt-Tuning Decision Transformer with Preference RankingFoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph Promptviz2viz: Prompt-driven stylized visualization generation using a diffusion modelText-driven Prompt Generation for Vision-Language Models in Federated LearningConsistency-guided Prompt Learning for Vision-Language ModelsLarge Language Model Prompt Chaining for Long Legal Document ClassificationPlum: Prompt Learning using Metaheuristic
Ready to Publish
Oct 31, 2023
Jun 28, 2024 06:54 PM
Efficient Federated Prompt Tuning for Black-box Large Pre-trained ModelsSPELL: Semantic Prompt Evolution based on a LLMMaatphor: Automated Variant Analysis for Prompt Injection AttacksImage-Object-Specific Prompt Learning for Few-Shot Class-Incremental Learning
Ready to Publish
Apr 4, 2023
Jun 28, 2024 06:48 PM
Efficient Federated Prompt Tuning for Black-box Large Pre-trained ModelsSPELL: Semantic Prompt Evolution based on a LLMPrompting Hard or Hardly Prompting: Prompt Inversion for Text-to-Image Diffusion ModelsFoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph PromptText-driven Prompt Generation for Vision-Language Models in Federated Learning
Ready to Publish
Oct 4, 2023
Jun 28, 2024 06:45 PM
Divide and Prompt: Chain of Thought Prompting for Text-to-SQLEfficient Federated Prompt Tuning for Black-box Large Pre-trained ModelsPrompt-based Node Feature Extractor for Few-shot Learning on Text-Attributed Graphsviz2viz: Prompt-driven stylized visualization generation using a diffusion modelRobust Safety Classifier for Large Language Models: Adversarial Prompt Shield
Ready to Publish
Oct 2, 2023
Jun 28, 2024 06:33 PM
Prompt-Tuning Decision Transformer with Preference RankingProgressive Visual Prompt Learning with Contrastive Feature Re-formationSoft-prompt Tuning for Large Language Models to Evaluate Biasviz2viz: Prompt-driven stylized visualization generation using a diffusion modelRobust Safety Classifier for Large Language Models: Adversarial Prompt Shield
Ready to Publish
Dec 12, 2023
Jun 28, 2024 06:28 PM
Prompt-Tuning Decision Transformer with Preference RankingProgressive Visual Prompt Learning with Contrastive Feature Re-formationSoft-prompt Tuning for Large Language Models to Evaluate BiasRobust Safety Classifier for Large Language Models: Adversarial Prompt Shield
Ready to Publish
Dec 19, 2023
Jun 28, 2024 06:22 PM
Progressive Visual Prompt Learning with Contrastive Feature Re-formationTesting LLMs on Code Generation with Varying Levels of Prompt SpecificityPrompting Hard or Hardly Prompting: Prompt Inversion for Text-to-Image Diffusion Modelsviz2viz: Prompt-driven stylized visualization generation using a diffusion model
Ready to Publish
Mar 5, 2024
Jun 28, 2024 06:13 PM
LLM Critics Help Catch LLM BugsBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalProgressive Visual Prompt Learning with Contrastive Feature Re-formationMaatphor: Automated Variant Analysis for Prompt Injection AttacksSPELL: Semantic Prompt Evolution based on a LLM
Ready to Publish
Dec 17, 2023
Jun 28, 2024 06:09 PM
Prompt-Tuning Decision Transformer with Preference RankingMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationTesting LLMs on Code Generation with Varying Levels of Prompt Specificity
Ready to Publish
Nov 10, 2023
Jun 28, 2024 06:04 PM
LLM Critics Help Catch LLM BugsPrompt-Tuning Decision Transformer with Preference RankingProgressive Visual Prompt Learning with Contrastive Feature Re-formationULTRA-DP: Unifying Graph Pre-training with Multi-task Graph Dual PromptPrompting Hard or Hardly Prompting: Prompt Inversion for Text-to-Image Diffusion Models
Ready to Publish
Apr 17, 2023
Jun 28, 2024 06:00 PM
An automatically discovered chain-of-thought prompt generalizes to novel models and datasetsVisual Prompt Based Personalized Federated LearningPromptTTS 2: Describing and Generating Voices with Text PromptTesting LLMs on Code Generation with Varying Levels of Prompt SpecificitySoft-prompt Tuning for Large Language Models to Evaluate BiasPrompting Hard or Hardly Prompting: Prompt Inversion for Text-to-Image Diffusion ModelsMaatphor: Automated Variant Analysis for Prompt Injection AttacksSPELL: Semantic Prompt Evolution based on a LLM
Ready to Publish
Jun 28, 2024
Featured
Jun 27, 2024 10:14 PM
Testing LLMs on Code Generation with Varying Levels of Prompt SpecificitySoft-prompt Tuning for Large Language Models to Evaluate BiasConsistency-guided Prompt Learning for Vision-Language ModelsFrom LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
Ready to Publish
Mar 26, 2024
Jun 27, 2024 06:49 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationPrompt Algebra for Task CompositionPrompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native Services
Ready to Publish
Jun 18, 2024
Jun 27, 2024 06:44 PM
Prompt Algebra for Task CompositionLast One Standing: A Comparative Analysis of Security and Privacy of Soft Prompt Tuning, LoRA, and In-Context LearningPrompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native Services
Ready to Publish
Oct 17, 2023
Jun 27, 2024 06:38 PM
Prompt-Tuning Decision Transformer with Preference RankingBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalPrompt Algebra for Task CompositionLMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual Recognition
Ready to Publish
Jun 4, 2023
Jun 27, 2024 06:18 PM
BIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalPrompt-Tuning Decision Transformer with Preference RankingPrompt Algebra for Task CompositionLMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual RecognitionSD4Match: Learning to Prompt Stable Diffusion Model for Semantic Matching
Ready to Publish
Nov 3, 2023
Jun 27, 2024 06:14 PM
Prompt-Tuning Decision Transformer with Preference RankingMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationSafeguarding Crowdsourcing Surveys from ChatGPT with Prompt Injection
Ready to Publish
Jun 15, 2023
Jun 27, 2024 06:03 PM
BIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationPrompt Algebra for Task CompositionAntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image Detectors
Ready to Publish
Jun 1, 2023
Jun 27, 2024 05:57 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationPrompt-Tuning Decision Transformer with Preference RankingAdversarial Prompt Tuning for Vision-Language ModelsSafeguarding Crowdsourcing Surveys from ChatGPT with Prompt InjectionPrompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native ServicesLast One Standing: A Comparative Analysis of Security and Privacy of Soft Prompt Tuning, LoRA, and In-Context LearningLMPT: Prompt Tuning with Class-Specific Embedding Loss for Long-tailed Multi-Label Visual RecognitionSD4Match: Learning to Prompt Stable Diffusion Model for Semantic Matching
Ready to Publish
May 16, 2023
Jun 27, 2024 05:52 PM
Prompt-Tuning Decision Transformer with Preference RankingBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalPromise: Prompt-driven 3D Medical Image Segmentation Using Pretrained Image Foundation ModelsPrompt Algebra for Task CompositionAntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image DetectorsPrompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native ServicesLast One Standing: A Comparative Analysis of Security and Privacy of Soft Prompt Tuning, LoRA, and In-Context LearningTesting LLMs on Code Generation with Varying Levels of Prompt SpecificityULTRA-DP: Unifying Graph Pre-training with Multi-task Graph Dual PromptMaatphor: Automated Variant Analysis for Prompt Injection AttacksSPELL: Semantic Prompt Evolution based on a LLMFoodGPT: A Large Language Model in Food Testing Domain with Incremental Pre-training and Knowledge Graph Prompt
Ready to Publish
Dec 25, 2023
Jun 27, 2024 05:48 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationPromise: Prompt-driven 3D Medical Image Segmentation Using Pretrained Image Foundation ModelsAutoHint: Automatic Prompt Optimization with Hint GenerationPrompt Algebra for Task Composition
Ready to Publish
Nov 13, 2023
Jun 27, 2024 05:45 PM
Promise: Prompt-driven 3D Medical Image Segmentation Using Pretrained Image Foundation ModelsMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationPrompt-In-Prompt Learning for Universal Image RestorationAdversarial Prompt Tuning for Vision-Language ModelsPrompt-Tuning Decision Transformer with Preference Ranking
Ready to Publish
Aug 8, 2023
Jun 26, 2024 01:31 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationTCP:Textual-based Class-aware Prompt tuning for Visual-Language ModelBadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIPAdversarial Prompt Tuning for Vision-Language Models
Ready to Publish
Dec 8, 2023
Jun 26, 2024 01:24 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationTCP:Textual-based Class-aware Prompt tuning for Visual-Language ModelBadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIPPromise: Prompt-driven 3D Medical Image Segmentation Using Pretrained Image Foundation Models
Ready to Publish
Mar 13, 2024
Jun 26, 2024 01:18 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerSoft Prompt Tuning for Augmenting Dense Retrieval with Large Language ModelsPrompt-In-Prompt Learning for Universal Image RestorationAutoHint: Automatic Prompt Optimization with Hint Generation
Ready to Publish
Feb 18, 2024
Jun 26, 2024 01:10 PM
DP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationControlling Personality Style in Dialogue with Zero-Shot Prompt-Based Learning
Ready to Publish
Apr 18, 2023
Jun 26, 2024 01:06 PM
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerControlling Personality Style in Dialogue with Zero-Shot Prompt-Based LearningPrompt-Tuning Decision Transformer with Preference RankingSafeguarding Crowdsourcing Surveys from ChatGPT with Prompt InjectionPrompt Sapper: LLM-Empowered Software Engineering Infrastructure for AI-Native ServicesLast One Standing: A Comparative Analysis of Security and Privacy of Soft Prompt Tuning, LoRA, and In-Context LearningSoft-prompt Tuning for Large Language Models to Evaluate Bias
Ready to Publish
Mar 20, 2024
Jun 26, 2024 12:57 PM
Layout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringControlling Personality Style in Dialogue with Zero-Shot Prompt-Based LearningBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalToken-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual InformationTCP:Textual-based Class-aware Prompt tuning for Visual-Language ModelPrompt-In-Prompt Learning for Universal Image RestorationAutoHint: Automatic Prompt Optimization with Hint GenerationPromise: Prompt-driven 3D Medical Image Segmentation Using Pretrained Image Foundation ModelsAdversarial Prompt Tuning for Vision-Language ModelsPrompt Algebra for Task CompositionSafeguarding Crowdsourcing Surveys from ChatGPT with Prompt InjectionAntifakePrompt: Prompt-Tuned Vision-Language Models are Fake Image DetectorsSD4Match: Learning to Prompt Stable Diffusion Model for Semantic MatchingULTRA-DP: Unifying Graph Pre-training with Multi-task Graph Dual Prompt
Ready to Publish
Mar 22, 2024
Jun 26, 2024 12:50 PM
Prompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerLayout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringPrompt-In-Prompt Learning for Universal Image RestorationAutoHint: Automatic Prompt Optimization with Hint Generation
Ready to Publish
Jun 17, 2024
Jun 26, 2024 12:45 PM
Layout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringSoft Prompt Tuning for Augmenting Dense Retrieval with Large Language ModelsTCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
Ready to Publish
Feb 8, 2023
Jun 26, 2024 12:42 PM
Layout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringIntroducing Athina Prompt Management: A Powerful and Flexible Prompt Playground and CMSMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image RestorationBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalToken-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual Information
Ready to Publish
Sep 7, 2023
Jun 26, 2024 12:36 PM
Prompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerIntroducing Athina Prompt Management: A Powerful and Flexible Prompt Playground and CMSControlling Personality Style in Dialogue with Zero-Shot Prompt-Based LearningSoft Prompt Tuning for Augmenting Dense Retrieval with Large Language ModelsBadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIPMultimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
Ready to Publish
Jun 26, 2024
Featured
Jun 25, 2024 08:40 PM
Cookbook: How to set up Langchain tracing on Athina in 2 minutesEvaluating LLM Chatbot Conversations with Athina AIAthina AI: LLM Monitoring and Evaluation PlatformLayout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringControlling Personality Style in Dialogue with Zero-Shot Prompt-Based Learning
Ready to Publish
Apr 23, 2023
Jun 25, 2024 08:16 PM
Prompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPrompt-based Node Feature Extractor for Few-shot Learning on Text-Attributed GraphsPrompt Middleware: Mapping Prompts for Large Language Models to UI AffordancesEfficient Federated Prompt Tuning for Black-box Large Pre-trained Models
Ready to Publish
Sep 6, 2023
Jun 25, 2024 08:10 PM
LLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPrompt-based Node Feature Extractor for Few-shot Learning on Text-Attributed GraphsDivide and Prompt: Chain of Thought Prompting for Text-to-SQLEfficient Federated Prompt Tuning for Black-box Large Pre-trained Models
Ready to Publish
Mar 25, 2023
Jun 25, 2024 08:06 PM
LLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersPrompt-Guided Transformers for End-to-End Open-Vocabulary Object DetectionPrompt Tuning Large Language Models on Personalized Aspect Extraction for Recommendations
Ready to Publish
Jul 3, 2023
Jun 25, 2024 08:00 PM
LLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPractical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt CalibrationDivide and Prompt: Chain of Thought Prompting for Text-to-SQL
Ready to Publish
Jun 2, 2023
Jun 25, 2024 07:55 PM
DP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPrompt Tuning Large Language Models on Personalized Aspect Extraction for RecommendationsPrompt-Guided Transformers for End-to-End Open-Vocabulary Object Detection
Ready to Publish
Dec 12, 2023
Featured
Jun 25, 2024 04:52 PM
Prompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerExploring the Relationship between LLM Hallucinations and Prompt Linguistic Nuances: Readability, Formality, and ConcretenessPrompt Middleware: Mapping Prompts for Large Language Models to UI Affordances
Ready to Publish
Sep 20, 2023
Jun 25, 2024 04:48 PM
LLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersAn automatically discovered chain-of-thought prompt generalizes to novel models and datasetsExploring the Relationship between LLM Hallucinations and Prompt Linguistic Nuances: Readability, Formality, and ConcretenessPractical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt Calibration
Ready to Publish
Mar 7, 2024
Jun 25, 2024 04:43 PM
An automatically discovered chain-of-thought prompt generalizes to novel models and datasetsLLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersVisual Prompt Based Personalized Federated Learning
Ready to Publish
Aug 3, 2023
Jun 25, 2024 04:40 PM
An automatically discovered chain-of-thought prompt generalizes to novel models and datasetsLLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPromptTTS 2: Describing and Generating Voices with Text PromptQuery-Dependent Prompt Evaluation and Optimization with Offline Inverse RLExploring the Relationship between LLM Hallucinations and Prompt Linguistic Nuances: Readability, Formality, and ConcretenessProgressive Visual Prompt Learning with Contrastive Feature Re-formation
Ready to Publish
Mar 17, 2024
Jun 25, 2024 04:36 PM
DP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerVisual Prompt Based Personalized Federated LearningPrompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech RecognitionPractical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt CalibrationPrompt Tuning Large Language Models on Personalized Aspect Extraction for RecommendationsLayout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringBadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIPBIM-GPT: a Prompt-Based Virtual Assistant Framework for BIM Information RetrievalToken-Level Adversarial Prompt Detection Based on Perplexity Measures and Contextual InformationTCP:Textual-based Class-aware Prompt tuning for Visual-Language Model
Ready to Publish
Mar 15, 2023
Jun 24, 2024 12:39 PM
Generalized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsPromptTTS 2: Describing and Generating Voices with Text PromptPrompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech RecognitionDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt EngineerQuery-Dependent Prompt Evaluation and Optimization with Offline Inverse RLProgressive Visual Prompt Learning with Contrastive Feature Re-formation
Ready to Publish
Oct 12, 2023
Jun 24, 2024 12:35 PM
Prompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech RecognitionHD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion ModelsGeneralized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsVisual Prompt Based Personalized Federated LearningAn automatically discovered chain-of-thought prompt generalizes to novel models and datasetsProgressive Visual Prompt Learning with Contrastive Feature Re-formation
Ready to Publish
Feb 16, 2023
Jun 24, 2024 12:30 PM
HD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion ModelsEdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAMPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationPromptTTS 2: Describing and Generating Voices with Text PromptVisual Prompt Based Personalized Federated LearningDP-OPT: Make Large Language Model Your Privacy-Preserving Prompt Engineer
Ready to Publish
Oct 16, 2023
Jun 24, 2024 12:22 PM
Generalized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsHD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion ModelsPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationAn automatically discovered chain-of-thought prompt generalizes to novel models and datasetsPractical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt CalibrationPrompt Tuning Large Language Models on Personalized Aspect Extraction for RecommendationsPrompt Middleware: Mapping Prompts for Large Language Models to UI AffordancesPrompt-based Node Feature Extractor for Few-shot Learning on Text-Attributed GraphsDivide and Prompt: Chain of Thought Prompting for Text-to-SQLLayout and Task Aware Instruction Prompt for Zero-shot Document Image Question AnsweringBadCLIP: Trigger-Aware Prompt Learning for Backdoor Attacks on CLIP
Ready to Publish
Mar 18, 2024
Jun 24, 2024 12:14 PM
Generalized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsDePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuningPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPrompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech RecognitionPromptTTS 2: Describing and Generating Voices with Text Prompt
Ready to Publish
Dec 10, 2023
Jun 24, 2024 11:59 AM
Are Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced SanitizationEdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAMDePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuningHD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion ModelsPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPromptTTS 2: Describing and Generating Voices with Text PromptVisual Prompt Based Personalized Federated Learning
Ready to Publish
Dec 11, 2023
Jun 24, 2024 11:55 AM
EdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAMAre Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced SanitizationPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationGeneralized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsPrompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech Recognition
Ready to Publish
Dec 15, 2023
Jun 24, 2024 11:51 AM
PromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationAre Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced SanitizationProRes: Exploring Degradation-aware Visual Prompt for Universal Image RestorationAn automatically discovered chain-of-thought prompt generalizes to novel models and datasetsQuery-Dependent Prompt Evaluation and Optimization with Offline Inverse RLExploring the Relationship between LLM Hallucinations and Prompt Linguistic Nuances: Readability, Formality, and ConcretenessPrompt Middleware: Mapping Prompts for Large Language Models to UI AffordancesPrompt-Guided Transformers for End-to-End Open-Vocabulary Object DetectionPrompt-based Node Feature Extractor for Few-shot Learning on Text-Attributed Graphs
Ready to Publish
May 24, 2023
Jun 24, 2024 11:37 AM
TopicGPT: A Prompt-based Topic Modeling FrameworkPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationPrompt-tuning latent diffusion models for inverse problemsLLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersEdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAMGeneralized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on Graphs
Ready to Publish
Nov 28, 2023
Jun 24, 2024 11:27 AM
DePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuningPrompt-tuning latent diffusion models for inverse problemsPBNR: Prompt-based News Recommender SystemAre Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced SanitizationLLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly TransformersEdgeSAM: Prompt-In-the-Loop Distillation for On-Device Deployment of SAMHD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion ModelsPrompt Packer: Deceiving LLMs through Compositional Instruction with Hidden AttacksPrompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech Recognition
Ready to Publish
Feb 18, 2024
Jun 24, 2024 11:22 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsProRes: Exploring Degradation-aware Visual Prompt for Universal Image RestorationPrompt-tuning latent diffusion models for inverse problemsPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationGeneralized Graph Prompt: Toward a Unification of Pre-Training and Downstream Tasks on GraphsHD-Painter: High-Resolution and Prompt-Faithful Text-Guided Image Inpainting with Diffusion Models
Ready to Publish
Jun 23, 2023
Jun 24, 2024 11:18 AM
TopicGPT: A Prompt-based Topic Modeling FrameworkLanguage Prompt for Autonomous DrivingPrompt-tuning latent diffusion models for inverse problemsDePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuningLLMs Can Understand Encrypted Prompt: Towards Privacy-Computing Friendly Transformers
Ready to Publish
Oct 2, 2023
Jun 24, 2024 11:12 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationIgnore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking CompetitionProRes: Exploring Degradation-aware Visual Prompt for Universal Image RestorationDePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuningPromptCARE: Prompt Copyright Protection by Watermark Injection and VerificationAre Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced Sanitization
Ready to Publish
Apr 1, 2024
Jun 24, 2024 11:04 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionProRes: Exploring Degradation-aware Visual Prompt for Universal Image RestorationAre Chatbots Ready for Privacy-Sensitive Applications? An Investigation into Input Regurgitation and Prompt-Induced Sanitization
Ready to Publish
Apr 15, 2024
Jun 24, 2024 10:57 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsSegment Any Anomaly without Training via Hybrid Prompt RegularizationImageDream: Image-Prompt Multi-view Diffusion for 3D Generation
Ready to Publish
Mar 3, 2024
Jun 24, 2024 10:51 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionPrompt-tuning latent diffusion models for inverse problems
Ready to Publish
Apr 16, 2023
Jun 24, 2024 10:43 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsSegment Any Anomaly without Training via Hybrid Prompt RegularizationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionPromptCARE: Prompt Copyright Protection by Watermark Injection and Verification
Ready to Publish
May 25, 2024
Jun 24, 2024 10:39 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
Aug 10, 2023
Jun 24, 2024 10:32 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsSegment Any Anomaly without Training via Hybrid Prompt RegularizationImageDream: Image-Prompt Multi-view Diffusion for 3D Generation
Ready to Publish
May 23, 2024
Jun 24, 2024 10:26 AM
ImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLanguage Prompt for Autonomous DrivingLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
Jan 8, 2024
Jun 24, 2024 10:22 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and Beyond
Ready to Publish
Nov 17, 2023
Jun 24, 2024 10:16 AM
Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
Apr 2, 2024
Jun 24, 2024 10:07 AM
Segment Any Anomaly without Training via Hybrid Prompt RegularizationLanguage Prompt for Autonomous DrivingLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
Aug 20, 2023
Jun 24, 2024 10:01 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
Mar 8, 2024
Jun 24, 2024 09:54 AM
Benchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsLanguage Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationStyleDiffusion: Prompt-Embedding Inversion for Text-Based EditingYou Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic ContentPBNR: Prompt-based News Recommender SystemPrompt Stealing Attacks Against Text-to-Image Generation ModelsDePT: Decomposed Prompt Tuning for Parameter-Efficient Fine-tuning
Ready to Publish
Aug 15, 2023
Jun 23, 2024 01:08 PM
From Prompt Injections to SQL Injection Attacks: How Protected is Your LLM-Integrated Web Application?Pre-Training to Learn in ContextThe Web Can Be Your Oyster for Improving Large Language ModelsEnhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design Strategies
Ready to Publish
May 21, 2023
Jun 23, 2024 01:07 PM
The Web Can Be Your Oyster for Improving Large Language ModelsFrom Prompt Injections to SQL Injection Attacks: How Protected is Your LLM-Integrated Web Application?Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTSpeechPrompt v2: Prompt Tuning for Speech Classification TasksPrivacy-Preserving Prompt Tuning for Large Language Model ServicesNegative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models
Ready to Publish
Mar 1, 2023
Jun 23, 2024 01:07 PM
The Web Can Be Your Oyster for Improving Large Language ModelsSpeechPrompt v2: Prompt Tuning for Speech Classification TasksEnhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design StrategiesPrivacy-Preserving Prompt Tuning for Large Language Model Services
Ready to Publish
May 10, 2023
Jun 23, 2024 01:06 PM
SpeechPrompt v2: Prompt Tuning for Speech Classification TasksEnhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design StrategiesCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT
Ready to Publish
May 26, 2023
Jun 23, 2024 01:06 PM
SatLM: Satisfiability-Aided Language Models Using Declarative PromptingThe Web Can Be Your Oyster for Improving Large Language ModelsEnhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design Strategies
Ready to Publish
Sep 8, 2023
Jun 23, 2024 01:05 PM
Pre-Training to Learn in ContextPlan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language ModelsPrompt Injection: Different Attacks and Defensive TechniquesSegment Any Anomaly without Training via Hybrid Prompt RegularizationLLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language ModelsBenchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsTEMPO: Prompt-based Generative Pre-trained Transformer for Time Series ForecastingPrompt a Robot to Walk with Large Language ModelsJatmo: Prompt Injection Defense by Task-Specific FinetuningReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingAssessing Prompt Injection Risks in 200+ Custom GPTsIgnore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking CompetitionTopicGPT: A Prompt-based Topic Modeling FrameworkPrompt-tuning latent diffusion models for inverse problemsProRes: Exploring Degradation-aware Visual Prompt for Universal Image Restoration
Ready to Publish
Dec 2, 2023
Jun 23, 2024 01:05 PM
Chain-of-Verification Reduces Hallucination in Large Language ModelsEfficient Prompting via Dynamic In-Context LearningPre-Training to Learn in ContextRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondLLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language ModelsPromptbreeder: Self-Referential Self-Improvement Via Prompt EvolutionBenchmarking and Defending Against Indirect Prompt Injection Attacks on Large Language ModelsPrompt a Robot to Walk with Large Language ModelsJatmo: Prompt Injection Defense by Task-Specific FinetuningReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingYou Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic ContentAssessing Prompt Injection Risks in 200+ Custom GPTsIgnore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking CompetitionPrompt Stealing Attacks Against Text-to-Image Generation ModelsTopicGPT: A Prompt-based Topic Modeling FrameworkPrompt-tuning latent diffusion models for inverse problems
Ready to Publish
Oct 10, 2023
Jun 23, 2024 01:05 PM
Efficient Prompting via Dynamic In-Context LearningChain-of-Verification Reduces Hallucination in Large Language ModelsPre-Training to Learn in ContextRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondPromptbreeder: Self-Referential Self-Improvement Via Prompt EvolutionConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingPrompt Injection attack against LLM-integrated ApplicationsStyleDiffusion: Prompt-Embedding Inversion for Text-Based EditingTEMPO: Prompt-based Generative Pre-trained Transformer for Time Series ForecastingPrompt a Robot to Walk with Large Language ModelsReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingAssessing Prompt Injection Risks in 200+ Custom GPTsPBNR: Prompt-based News Recommender SystemIgnore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking CompetitionTopicGPT: A Prompt-based Topic Modeling Framework
Ready to Publish
May 18, 2023
Jun 23, 2024 01:04 PM
Language Prompt for Autonomous DrivingCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable PromptEfficient Prompting via Dynamic In-Context LearningTEMPO: Prompt-based Generative Pre-trained Transformer for Time Series ForecastingYou Only Prompt Once: On the Capabilities of Prompt Learning on Large Language Models to Tackle Toxic ContentPBNR: Prompt-based News Recommender SystemPrompt Stealing Attacks Against Text-to-Image Generation Models
Ready to Publish
Apr 26, 2023
Jun 23, 2024 01:04 PM
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationEfficient Prompting via Dynamic In-Context LearningPromptbreeder: Self-Referential Self-Improvement Via Prompt EvolutionConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignStyleDiffusion: Prompt-Embedding Inversion for Text-Based EditingJatmo: Prompt Injection Defense by Task-Specific Finetuning
Ready to Publish
Mar 4, 2024
Jun 23, 2024 01:03 PM
ImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLanguage Prompt for Autonomous DrivingEfficient Prompting via Dynamic In-Context LearningQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software Design
Ready to Publish
Sep 28, 2023
Jun 23, 2024 01:03 PM
Re-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software Design
Ready to Publish
Oct 17, 2023
Jun 23, 2024 01:02 PM
Promptbreeder: Self-Referential Self-Improvement Via Prompt EvolutionLLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language ModelsLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionPrompt Injection attack against LLM-integrated ApplicationsJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsTensor Trust: Interpretable Prompt Injection Attacks from an Online GameAnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly Detection
Ready to Publish
Mar 11, 2023
Jun 23, 2024 01:02 PM
Promptbreeder: Self-Referential Self-Improvement Via Prompt EvolutionLLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language ModelsRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondPrompt Injection attack against LLM-integrated ApplicationsJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsTensor Trust: Interpretable Prompt Injection Attacks from an Online GameAnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly DetectionBlack-Box Prompt Optimization: Aligning Large Language Models without Model Training
Ready to Publish
Mar 2, 2024
Jun 23, 2024 01:01 PM
ChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionJailbreaking ChatGPT via Prompt Engineering: An Empirical Study
Ready to Publish
Mar 10, 2024
Jun 23, 2024 01:01 PM
Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingPrompt Injection attack against LLM-integrated ApplicationsChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsTensor Trust: Interpretable Prompt Injection Attacks from an Online GameAnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly DetectionAn LLM can Fool Itself: A Prompt-Based Adversarial AttackPromptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code Generators
Ready to Publish
Aug 13, 2023
Jun 23, 2024 01:00 PM
Jailbreaking ChatGPT via Prompt Engineering: An Empirical StudyChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingPromptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code GeneratorsPromptAid: Prompt Exploration, Perturbation, Testing and Iteration using Visual Analytics for Large Language ModelsBlack-Box Prompt Optimization: Aligning Large Language Models without Model TrainingBoosted Prompt Ensembles for Large Language Models
Ready to Publish
Nov 2, 2023
Jun 23, 2024 01:00 PM
ChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingAn LLM can Fool Itself: A Prompt-Based Adversarial Attack
Ready to Publish
Mar 16, 2024
Jun 23, 2024 12:59 PM
ChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software DesignQuantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formattingJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyAn LLM can Fool Itself: A Prompt-Based Adversarial AttackPromptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code Generators
Ready to Publish
Oct 20, 2023
Jun 23, 2024 12:59 PM
AnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly DetectionTensor Trust: Interpretable Prompt Injection Attacks from an Online GameJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyPromptAid: Prompt Exploration, Perturbation, Testing and Iteration using Visual Analytics for Large Language Models
Ready to Publish
Jul 31, 2023
Jun 23, 2024 12:58 PM
AnomalyCLIP: Object-agnostic Prompt Learning for Zero-shot Anomaly DetectionJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsPromptAid: Prompt Exploration, Perturbation, Testing and Iteration using Visual Analytics for Large Language ModelsBlack-Box Prompt Optimization: Aligning Large Language Models without Model TrainingBoosted Prompt Ensembles for Large Language Models
Ready to Publish
Apr 8, 2023
Jun 23, 2024 12:58 PM
Promptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code GeneratorsIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsAn LLM can Fool Itself: A Prompt-Based Adversarial AttackBoosted Prompt Ensembles for Large Language Models
Ready to Publish
Nov 8, 2023
Jun 23, 2024 12:57 PM
Promptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code GeneratorsIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsChatGPT Prompt Patterns for Improving Code Quality, Refactoring, Requirements Elicitation, and Software Design
Ready to Publish
Apr 12, 2023
Jun 23, 2024 12:56 PM
Promptly: Using Prompt Problems to Teach Learners How to Effectively Utilize AI Code GeneratorsIP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion ModelsPromptAid: Prompt Exploration, Perturbation, Testing and Iteration using Visual Analytics for Large Language Models
Ready to Publish
Feb 27, 2024
Jun 23, 2024 11:56 AM
LongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondPromptbreeder: Self-Referential Self-Improvement Via Prompt Evolution
Ready to Publish
Mar 2, 2023
Jun 22, 2024 01:19 PM
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataMultitask Prompt Tuning Enables Parameter-Efficient Transfer LearningChain of Hindsight Aligns Language Models with FeedbackLanguage Is Not All You Need: Aligning Perception with Language ModelsBounding the Capabilities of Large Language Models in Open Text Generation with Prompt ConstraintsA-la-carte Prompt Tuning (APT): Combining Distinct Data Via Composable PromptingEnhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design StrategiesPrivacy-Preserving Prompt Tuning for Large Language Model Services
Ready to Publish
Sep 20, 2023
Jun 22, 2024 12:10 PM
Deficiency of Large Language Models in Finance: An Empirical Examination of HallucinationCYBERSECEVAL 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language ModelsImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
May 6, 2023
Jun 22, 2024 12:04 PM
Can We Edit Factual Knowledge by In-Context Learning?A Bibliometric Review of Large Language Models Research from 2017 to 2023Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language ModelsMeta-in-context learning in large language modelsEfficient Prompting via Dynamic In-Context LearningLanguage Prompt for Autonomous Driving
Ready to Publish
May 17, 2023
Jun 22, 2024 12:04 PM
Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsMeta-in-context learning in large language modelsCan We Edit Factual Knowledge by In-Context Learning?TELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksFlatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencySegment Any Anomaly without Training via Hybrid Prompt Regularization
Ready to Publish
May 18, 2023
Jun 22, 2024 12:04 PM
TELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksTreePrompt: Learning to Compose Tree Prompts for Explainable Visual GroundingPlan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language ModelsChain-of-Symbol Prompting Elicits Planning in Large Langauge ModelsWhat In-Context Learning "Learns" In-Context: Disentangling Task Recognition and Task LearningReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt CompressionSegment Any Anomaly without Training via Hybrid Prompt RegularizationRe-imagine the Negative Prompt Algorithm: Transform 2D Diffusion into 3D, alleviate Janus problem and BeyondLLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models
Ready to Publish
May 22, 2023
Jun 22, 2024 12:03 PM
Reinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondExplaining Emergent In-Context Learning as Kernel RegressionInteractive Natural Language ProcessingMeta-in-context learning in large language modelsLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex Tasks
Ready to Publish
May 22, 2023
Jun 22, 2024 12:03 PM
Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language ModelsExplaining Emergent In-Context Learning as Kernel RegressionMeta-in-context learning in large language modelsCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable PromptTreePrompt: Learning to Compose Tree Prompts for Explainable Visual Grounding
Ready to Publish
May 19, 2023
Jun 22, 2024 12:02 PM
Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsMeta-in-context learning in large language modelsLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsEfficient Prompting via Dynamic In-Context LearningThe Web Can Be Your Oyster for Improving Large Language ModelsFlatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencyChain-of-Symbol Prompting Elicits Planning in Large Langauge Models
Ready to Publish
May 16, 2023
Jun 22, 2024 12:01 PM
Flatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencyZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMsEfficient Prompting via Dynamic In-Context Learning
Ready to Publish
May 17, 2023
Jun 22, 2024 12:00 PM
Efficient Prompting via Dynamic In-Context LearningZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMsTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksBoosted Prompt Ensembles for Large Language Models
Ready to Publish
May 16, 2023
Jun 22, 2024 12:00 PM
Flatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencyZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMsTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksPre-Training to Learn in ContextBoosted Prompt Ensembles for Large Language ModelsNegative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models
Ready to Publish
May 19, 2023
Jun 22, 2024 11:59 AM
Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsExplaining Emergent In-Context Learning as Kernel RegressionLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable PromptTreePrompt: Learning to Compose Tree Prompts for Explainable Visual GroundingTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksThe Web Can Be Your Oyster for Improving Large Language Models
Ready to Publish
May 19, 2023
Jun 22, 2024 11:58 AM
Explaining Emergent In-Context Learning as Kernel RegressionLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable PromptEfficient Prompting via Dynamic In-Context LearningThe Web Can Be Your Oyster for Improving Large Language ModelsFlatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencyReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingSatLM: Satisfiability-Aided Language Models Using Declarative PromptingPre-Training to Learn in Context
Ready to Publish
May 18, 2023
Jun 22, 2024 11:58 AM
TreePrompt: Learning to Compose Tree Prompts for Explainable Visual GroundingLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksChain-of-Symbol Prompting Elicits Planning in Large Langauge ModelsFrom Prompt Injections to SQL Injection Attacks: How Protected is Your LLM-Integrated Web Application?Enhancing Few-shot Text-to-SQL Capabilities of Large Language Models: A Study on Prompt Design StrategiesSpeechPrompt v2: Prompt Tuning for Speech Classification TasksNegative-prompt Inversion: Fast Image Inversion for Editing with Text-guided Diffusion Models
Ready to Publish
May 18, 2023
Jun 22, 2024 11:57 AM
TELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable PromptTreePrompt: Learning to Compose Tree Prompts for Explainable Visual GroundingWhat In-Context Learning "Learns" In-Context: Disentangling Task Recognition and Task LearningSatLM: Satisfiability-Aided Language Models Using Declarative PromptingBoosted Prompt Ensembles for Large Language Models
Ready to Publish
May 18, 2023
Jun 22, 2024 11:57 AM
Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMsWhat In-Context Learning "Learns" In-Context: Disentangling Task Recognition and Task LearningReprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingSatLM: Satisfiability-Aided Language Models Using Declarative PromptingPre-Training to Learn in Context
Ready to Publish
May 16, 2023
Jun 22, 2024 11:56 AM
ZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMsTELeR: A General Taxonomy of LLM Prompts for Benchmarking Complex TasksSatLM: Satisfiability-Aided Language Models Using Declarative PromptingFrom Prompt Injections to SQL Injection Attacks: How Protected is Your LLM-Integrated Web Application?Language Prompt for Autonomous DrivingImageDream: Image-Prompt Multi-view Diffusion for 3D GenerationLongLLMLingua: Accelerating and Enhancing LLMs in Long Context Scenarios via Prompt Compression
Ready to Publish
May 17, 2023
Jun 22, 2024 11:56 AM
Efficient Prompting via Dynamic In-Context LearningThe Web Can Be Your Oyster for Improving Large Language ModelsTreePrompt: Learning to Compose Tree Prompts for Explainable Visual Grounding
Ready to Publish
Apr 10, 2024
Jun 22, 2024 11:53 AM
Universal and Transferable Adversarial Attacks on Aligned Language ModelsSelf-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation
Ready to Publish
Apr 6, 2024
Featured
Jun 22, 2024 11:53 AM
Breaking Down the Defenses: A Comparative Survey of Attacks on Large Language ModelsEver: Mitigating Hallucination in Large Language Models through Real-Time Verification and RectificationSemi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model Reasoning
Summary of Anthropic Research on Many-Shot Jailbreaking
Summary of Anthropic Research on Many-Shot Jailbreaking
Ready to Publish
Apr 5, 2024
Featured
Jun 22, 2024 11:52 AM
Prompt Injection: Different Attacks and Defensive TechniquesMany-Shot Jailbreaking (Anthropic Research)Post-Semantic-Thinking: A Robust Strategy to Distill Reasoning Capacity from Large Language ModelsEver: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
Keeping Large Language Models Safe: What You Need to Know
Summary of Research Paper: Breaking Down the Defenses: A Comparative Survey of Attacks on Large Language Models
Ready to Publish
Apr 16, 2024
Featured
Jun 22, 2024 11:52 AM
Universal and Transferable Adversarial Attacks on Aligned Language ModelsDetect LLM Hallucinations in CI / CD: Evaluate your RAG pipeline using GitHub Actions + Athina / RagasPrompt Injection: Different Attacks and Defensive TechniquesWizardLM: Empowering Large Language Models to Follow Complex InstructionsEntGPT: Linking Generative Large Language Models with Knowledge Bases
Ready to Publish
Apr 14, 2024
Jun 22, 2024 11:51 AM
Prompt Injection: Different Attacks and Defensive TechniquesAI Safety: Necessary, but insufficient and possibly problematicFrom Noise to Clarity: Unraveling the Adversarial Suffix of Large Language Model Attacks via Translation of Text EmbeddingsMistral 7B: Foundation Model Research Paper SummaryWizardLM: Empowering Large Language Models to Follow Complex InstructionsEntGPT: Linking Generative Large Language Models with Knowledge Bases
Ready to Publish
Apr 18, 2024
Jun 22, 2024 11:51 AM
Prompt Injection: Different Attacks and Defensive TechniquesHow to Use a Custom Grading Criteria to Evaluate LLM Responses (LLM-as-a-Judge)Mistral 7B: Foundation Model Research Paper SummaryChain-of-Verification Reduces Hallucination in Large Language Models
Ready to Publish
Feb 23, 2023
Jun 22, 2024 11:50 AM
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersLanguage Is Not All You Need: Aligning Perception with Language ModelsEvoPrompting: Language Models for Code-Level Neural Architecture SearchA Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPT
Ready to Publish
May 26, 2023
Jun 22, 2024 11:47 AM
Exploring LLM-based Agents for Root Cause AnalysisAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled DataFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationTool Learning with Foundation ModelsTemporal evolution of depolarization and magnetic field of FRB 20201124AUniversality and Limitations of Prompt TuningMultiTool-CoT: GPT-3 Can Use Multiple External Tools with Chain of Thought PromptingPEARL: Prompting Large Language Models to Plan and Execute Actions Over Long DocumentsReasoning with Language Model is Planning with World ModelBetter Zero-Shot Reasoning with Self-Adaptive PromptingInteractive Natural Language ProcessingCan We Edit Factual Knowledge by In-Context Learning?
Ready to Publish
Apr 26, 2023
Jun 22, 2024 11:47 AM
Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondExploring LLM-based Agents for Root Cause AnalysisAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled DataOne Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraNatural Language Reasoning, A SurveyAugmented Language Models: a SurveyA Survey on In-context LearningTowards Reasoning in Large Language Models: A SurveyHierarchical Prompting Assists Large Language Model on Web NavigationZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMs
Ready to Publish
Apr 17, 2023
Jun 22, 2024 11:46 AM
Few-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondTool Learning with Foundation Models
Ready to Publish
Apr 3, 2023
Jun 22, 2024 11:46 AM
Reinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondOne Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyFocused Prefix Tuning for Controllable Text GenerationLess Likely Brainstorming: Using Language Models to Generate Alternative HypothesesPEARL: Prompting Large Language Models to Plan and Execute Actions Over Long DocumentsHierarchical Prompting Assists Large Language Model on Web NavigationCan We Edit Factual Knowledge by In-Context Learning?Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models
Ready to Publish
Dec 31, 2022
Jun 22, 2024 11:45 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondAugmented Language Models: a SurveyReasoning with Language Model Prompting: A Survey
Ready to Publish
Sep 13, 2023
Jun 22, 2024 11:45 AM
Reasoning with Language Model Prompting: A SurveyTowards Reasoning in Large Language Models: A SurveyFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
May 31, 2023
Jun 22, 2024 11:44 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraReasoning with Language Model Prompting: A SurveyReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and Beyond
Ready to Publish
May 22, 2023
Jun 22, 2024 11:44 AM
Few-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyCan We Edit Factual Knowledge by In-Context Learning?Explaining Emergent In-Context Learning as Kernel Regression
Ready to Publish
Apr 7, 2023
Jun 22, 2024 11:42 AM
Revisiting Automated Prompting: Are We Actually Doing Better?A Comprehensive Survey on Instruction FollowingGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceDynamic Prompting: A Unified Framework for Prompt TuningEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited Data
Ready to Publish
Mar 20, 2023
Jun 22, 2024 11:42 AM
Boosted Prompt Ensembles for Large Language ModelsGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningWhy think step by step? Reasoning emerges from the locality of experience
Ready to Publish
Mar 30, 2023
Jun 22, 2024 11:41 AM
Boosted Prompt Ensembles for Large Language ModelsGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningWhy think step by step? Reasoning emerges from the locality of experience
Ready to Publish
Mar 24, 2023
Jun 22, 2024 11:41 AM
A Comprehensive Survey on Instruction FollowingRevisiting Automated Prompting: Are We Actually Doing Better?Global Prompt Cell: A Portable Control Module for Effective Prompt TuningContext-faithful Prompting for Large Language ModelsStructure Pretraining and Prompt Tuning for Knowledge Graph TransferCoTEVer: Chain of Thought Prompting Annotation Toolkit for Explanation VerificationLarger language models do in-context learning differently
Ready to Publish
Mar 23, 2023
Jun 22, 2024 11:41 AM
Boosted Prompt Ensembles for Large Language ModelsFairness-guided Few-shot Prompting for Large Language ModelsVisual-Language Prompt Tuning with Knowledge-guided Context OptimizationContext-faithful Prompting for Large Language ModelsStructure Pretraining and Prompt Tuning for Knowledge Graph TransferCoTEVer: Chain of Thought Prompting Annotation Toolkit for Explanation VerificationLarger language models do in-context learning differently
Ready to Publish
Mar 20, 2023
Jun 22, 2024 11:40 AM
Boosted Prompt Ensembles for Large Language ModelsFairness-guided Few-shot Prompting for Large Language ModelsNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor Inference
Ready to Publish
Mar 7, 2023
Jun 22, 2024 11:40 AM
Boosted Prompt Ensembles for Large Language ModelsFairness-guided Few-shot Prompting for Large Language ModelsNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceDynamic Prompting: A Unified Framework for Prompt TuningART: Automatic multi-step reasoning and tool-use for large language models
Ready to Publish
Mar 7, 2023
Jun 22, 2024 11:39 AM
Boosted Prompt Ensembles for Large Language ModelsFairness-guided Few-shot Prompting for Large Language ModelsNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceOpenICL: An Open-Source Framework for In-context LearningAlphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Ready to Publish
Mar 6, 2023
Jun 22, 2024 11:36 AM
Larger language models do in-context learning differentlyBoosted Prompt Ensembles for Large Language ModelsStructure Pretraining and Prompt Tuning for Knowledge Graph TransferART: Automatic multi-step reasoning and tool-use for large language models
Ready to Publish
Mar 1, 2023
Jun 22, 2024 11:35 AM
Prompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersMultitask Prompt Tuning Enables Parameter-Efficient Transfer LearningMixture of Soft Prompts for Controllable Data GenerationEvoPrompting: Language Models for Code-Level Neural Architecture SearchChain of Hindsight Aligns Language Models with FeedbackLanguage Is Not All You Need: Aligning Perception with Language Models
Ready to Publish
Feb 6, 2023
Jun 22, 2024 11:34 AM
Chain of Hindsight Aligns Language Models with FeedbackCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTHow Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding TasksHow Does In-Context Learning Help Prompt Tuning?Scalable Prompt Generation for Semi-supervised Learning with Language Models
Ready to Publish
Feb 23, 2023
Jun 22, 2024 11:34 AM
Language Is Not All You Need: Aligning Perception with Language ModelsPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersEvoPrompting: Language Models for Code-Level Neural Architecture SearchA Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPTGuiding Large Language Models via Directional Stimulus Prompting
Ready to Publish
Feb 17, 2023
Jun 22, 2024 11:31 AM
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataA-la-carte Prompt Tuning (APT): Combining Distinct Data Via Composable PromptingGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksThe Capacity for Moral Self-Correction in Large Language Models
Ready to Publish
Apr 11, 2024
Jun 22, 2024 11:30 AM
LLM evaluation too expensive? Here's how we solve this.IterAlign: Iterative Constitutional Alignment of Large Language ModelsReflexion: Language Agents with Verbal Reinforcement LearningSemi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model Reasoning
Ready to Publish
Nov 14, 2023
Featured
Jun 22, 2024 11:29 AM
How to Evaluate AI Chats Using Conversation Coherence EvaluatorChain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous SourcesKnowGPT: Knowledge Injection for Large Language ModelsSiren's Song in the AI Ocean: A Survey on Hallucination in Large Language ModelsActive Retrieval Augmented Generation
Ready to Publish
Jun 1, 2023
Jun 22, 2024 11:29 AM
Towards Reasoning in Large Language Models: A SurveyA Bibliometric Review of Large Language Models Research from 2017 to 2023Reasoning with Language Model Prompting: A Survey
Ready to Publish
Apr 12, 2023
Jun 22, 2024 11:28 AM
A Comprehensive Survey on Instruction FollowingExploring LLM-based Agents for Root Cause AnalysisPrompt Design and Engineering: Introduction and Advanced MethodsWhy think step by step? Reasoning emerges from the locality of experienceRevisiting Automated Prompting: Are We Actually Doing Better?REFINER: Reasoning Feedback on Intermediate RepresentationsReflexion: Language Agents with Verbal Reinforcement LearningCAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietySelf-Refine: Iterative Refinement with Self-FeedbackNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceVisual-Language Prompt Tuning with Knowledge-guided Context Optimization
Ready to Publish
Mar 23, 2023
Jun 22, 2024 11:28 AM
Global Prompt Cell: A Portable Control Module for Effective Prompt TuningA Comprehensive Survey on Instruction FollowingWhy think step by step? Reasoning emerges from the locality of experienceFairness-guided Few-shot Prompting for Large Language Models
Ready to Publish
Mar 6, 2023
Jun 22, 2024 11:28 AM
Multitask Prompt Tuning Enables Parameter-Efficient Transfer LearningART: Automatic multi-step reasoning and tool-use for large language modelsDynamic Prompting: A Unified Framework for Prompt TuningMixture of Soft Prompts for Controllable Data GenerationHow Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding TasksCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERT
Ready to Publish
Mar 5, 2023
Jun 22, 2024 11:27 AM
A Comprehensive Survey on Instruction FollowingRevisiting Automated Prompting: Are We Actually Doing Better?Effectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataMixture of Soft Prompts for Controllable Data GenerationPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTEvoPrompting: Language Models for Code-Level Neural Architecture SearchScalable Prompt Generation for Semi-supervised Learning with Language ModelsBounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints
Ready to Publish
Feb 22, 2023
Jun 22, 2024 11:26 AM
Guiding Large Language Models via Directional Stimulus PromptingA Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPTChain of Hindsight Aligns Language Models with FeedbackScalable Prompt Generation for Semi-supervised Learning with Language ModelsGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksThe Capacity for Moral Self-Correction in Large Language Models
Ready to Publish
Feb 15, 2023
Jun 22, 2024 11:26 AM
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt ConstraintsCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTA Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPTGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksSwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource Domains
Ready to Publish
Apr 12, 2024
Jun 22, 2024 11:24 AM
IterAlign: Iterative Constitutional Alignment of Large Language ModelsBreaking Down the Defenses: A Comparative Survey of Attacks on Large Language Models10 Rules for LLM Evaluation: What we learned after a year of buildingHow to evaluate your Llama Index query engine using Ragas evals + Athina AI
Ready to Publish
Oct 9, 2023
Jun 22, 2024 11:24 AM
Exploring LLM-based Agents for Root Cause AnalysisInvestigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction FollowingModel-tuning Via Prompts Makes NLP Models Adversarially RobustJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyTool Learning with Foundation ModelsOne Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraA Bibliometric Review of Large Language Models Research from 2017 to 2023Natural Language Reasoning, A SurveyWalking Down the Memory Maze: Beyond Context Limit through Interactive ReadingFrom Sparse to Dense: GPT-4 Summarization with Chain of Density PromptingExploring Lottery Prompts for Pre-trained Language ModelsLet's Verify Step by StepPEARL: Prompting Large Language Models to Plan and Execute Actions Over Long DocumentsReasoning with Language Model is Planning with World ModelBetter Zero-Shot Reasoning with Self-Adaptive PromptingInteractive Natural Language ProcessingExplaining Emergent In-Context Learning as Kernel RegressionZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMs
Ready to Publish
Mar 26, 2023
Jun 22, 2024 11:24 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondReasoning with Language Model Prompting: A Survey
Ready to Publish
Feb 15, 2023
Jun 22, 2024 11:23 AM
Augmented Language Models: a SurveyOne Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondA Survey on In-context LearningTowards Reasoning in Large Language Models: A SurveyEmergent Abilities of Large Language ModelsPre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
Ready to Publish
Dec 20, 2022
Jun 22, 2024 11:23 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraAugmented Language Models: a SurveyHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondEmergent Abilities of Large Language ModelsA Taxonomy of Prompt Modifiers for Text-To-Image GenerationPre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language ProcessingWalking Down the Memory Maze: Beyond Context Limit through Interactive ReadingTemporal evolution of depolarization and magnetic field of FRB 20201124AChain-of-Verification Reduces Hallucination in Large Language ModelsFrom Sparse to Dense: GPT-4 Summarization with Chain of Density PromptingGraph of Thoughts: Solving Elaborate Problems with Large Language ModelsFocused Prefix Tuning for Controllable Text GenerationExploring Lottery Prompts for Pre-trained Language Models
Ready to Publish
Dec 19, 2022
Jun 22, 2024 11:23 AM
A Survey on In-context LearningNatural Language Reasoning, A SurveyEmergent Abilities of Large Language ModelsA Taxonomy of Prompt Modifiers for Text-To-Image GenerationWalking Down the Memory Maze: Beyond Context Limit through Interactive ReadingTemporal evolution of depolarization and magnetic field of FRB 20201124AChain-of-Verification Reduces Hallucination in Large Language ModelsGraph of Thoughts: Solving Elaborate Problems with Large Language ModelsFocused Prefix Tuning for Controllable Text GenerationExploring Lottery Prompts for Pre-trained Language ModelsLess Likely Brainstorming: Using Language Models to Generate Alternative HypothesesLet's Verify Step by Step
Ready to Publish
Jun 15, 2022
Jun 22, 2024 11:22 AM
Reasoning with Language Model Prompting: A SurveyTowards Reasoning in Large Language Models: A SurveyAugmented Language Models: a SurveyA Taxonomy of Prompt Modifiers for Text-To-Image Generation
Ready to Publish
Oct 8, 2023
Jun 22, 2024 11:22 AM
Reasoning with Language Model Prompting: A SurveyTowards Reasoning in Large Language Models: A SurveyReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and Beyond
Ready to Publish
May 30, 2023
Jun 22, 2024 11:22 AM
A Bibliometric Review of Large Language Models Research from 2017 to 2023Reasoning with Language Model Prompting: A SurveyGraph of Thoughts: Solving Elaborate Problems with Large Language Models
Ready to Publish
May 26, 2023
Jun 22, 2024 11:21 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraGraph of Thoughts: Solving Elaborate Problems with Large Language ModelsFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
May 24, 2023
Jun 22, 2024 11:21 AM
Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
May 22, 2023
Jun 22, 2024 11:20 AM
Few-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationA Bibliometric Review of Large Language Models Research from 2017 to 2023Interactive Natural Language ProcessingPlan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language ModelsCompress, Then Prompt: Improving Accuracy-Efficiency Trade-off of LLM Inference with Transferable Prompt
Ready to Publish
Apr 7, 2023
Jun 22, 2024 11:18 AM
Why think step by step? Reasoning emerges from the locality of experienceGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningEvaluating LLM Chatbot Conversations with Athina AIREFINER: Reasoning Feedback on Intermediate RepresentationsReflexion: Language Agents with Verbal Reinforcement LearningCAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietySelf-Refine: Iterative Refinement with Self-FeedbackVisual-Language Prompt Tuning with Knowledge-guided Context Optimization
Ready to Publish
Apr 4, 2023
Jun 22, 2024 11:18 AM
Global Prompt Cell: A Portable Control Module for Effective Prompt TuningBoosted Prompt Ensembles for Large Language ModelsWhy think step by step? Reasoning emerges from the locality of experience
Ready to Publish
Mar 16, 2023
Jun 22, 2024 11:18 AM
Dynamic Prompting: A Unified Framework for Prompt TuningCoTEVer: Chain of Thought Prompting Annotation Toolkit for Explanation VerificationOpenICL: An Open-Source Framework for In-context LearningMultitask Prompt Tuning Enables Parameter-Efficient Transfer LearningPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot Learners
Ready to Publish
Apr 4, 2024
Jun 22, 2024 11:17 AM
Universal and Transferable Adversarial Attacks on Aligned Language ModelsPrompt Injection: Different Attacks and Defensive TechniquesText Summarization: LLM Failure Cases and Detection MethodsCYBERSECEVAL 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language Models
Ready to Publish
Apr 4, 2023
Jun 22, 2024 11:16 AM
Reinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondOne Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraA Bibliometric Review of Large Language Models Research from 2017 to 2023Natural Language Reasoning, A SurveyAugmented Language Models: a SurveyA Survey on In-context LearningTowards Reasoning in Large Language Models: A SurveyChain-of-Verification Reduces Hallucination in Large Language ModelsLet's Verify Step by StepUniversality and Limitations of Prompt TuningMultiTool-CoT: GPT-3 Can Use Multiple External Tools with Chain of Thought Prompting
Ready to Publish
May 30, 2023
Jun 22, 2024 11:16 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraGraph of Thoughts: Solving Elaborate Problems with Large Language ModelsFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
Mar 3, 2023
Jun 22, 2024 11:16 AM
Effectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataMixture of Soft Prompts for Controllable Data GenerationART: Automatic multi-step reasoning and tool-use for large language modelsHow Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding TasksCan ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTEvoPrompting: Language Models for Code-Level Neural Architecture SearchLanguage Is Not All You Need: Aligning Perception with Language ModelsActive Prompting with Chain-of-Thought for Large Language ModelsNot what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt InjectionBounding the Capabilities of Large Language Models in Open Text Generation with Prompt Constraints
Ready to Publish
Apr 17, 2024
Jun 22, 2024 11:15 AM
From Noise to Clarity: Unraveling the Adversarial Suffix of Large Language Model Attacks via Translation of Text EmbeddingsUniversal and Transferable Adversarial Attacks on Aligned Language ModelsPrompt Injection: Different Attacks and Defensive TechniquesSearch-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive Tasks
Ready to Publish
Mar 2, 2023
Jun 22, 2024 11:14 AM
Multitask Prompt Tuning Enables Parameter-Efficient Transfer LearningEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataMixture of Soft Prompts for Controllable Data GenerationPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersHow Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding Tasks
Ready to Publish
Dec 11, 2023
Jun 22, 2024 11:13 AM
KnowGPT: Knowledge Injection for Large Language ModelsEntGPT: Linking Generative Large Language Models with Knowledge BasesFine-tuning Language Models for Factuality
Ready to Publish
May 11, 2023
Jun 22, 2024 11:12 AM
Fine-tuning Language Models for FactualitySearch-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive TasksAutoHall: Automated Hallucination Dataset Generation for Large Language ModelsPrompt Design and Engineering: Introduction and Advanced MethodsEnhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4A Comprehensive Survey on Instruction Following
Ready to Publish
Nov 13, 2023
Jun 22, 2024 11:12 AM
AutoHall: Automated Hallucination Dataset Generation for Large Language ModelsProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsDeficiency of Large Language Models in Finance: An Empirical Examination of HallucinationEnhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4A Comprehensive Survey on Instruction Following
Ready to Publish
Mar 3, 2023
Jun 22, 2024 11:11 AM
Fairness-guided Few-shot Prompting for Large Language ModelsBoosted Prompt Ensembles for Large Language ModelsNN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceOpenICL: An Open-Source Framework for In-context Learning
Ready to Publish
Feb 27, 2023
Jun 22, 2024 11:11 AM
Can ChatGPT Understand Too? A Comparative Study on ChatGPT and Fine-tuned BERTHow Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding TasksPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersActive Prompting with Chain-of-Thought for Large Language ModelsNot what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt InjectionA Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPTGuiding Large Language Models via Directional Stimulus Prompting
Ready to Publish
Apr 15, 2024
Featured
The EVER (Real-Time Verification and Rectification) framework is designed to dynamically mitigate hallucinations during text generation by ensuring the accuracy and trustworthiness of each sentence before proceeding.
Jun 22, 2024 11:10 AM
Many-Shot Jailbreaking (Anthropic Research)Prompt Injection: Different Attacks and Defensive TechniquesBreaking Down the Defenses: A Comparative Survey of Attacks on Large Language ModelsSemi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model ReasoningSearch-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive Tasks
Ready to Publish
Feb 4, 2024
Jun 22, 2024 11:09 AM
Factuality of Large Language Models in the Year 2024EntGPT: Linking Generative Large Language Models with Knowledge BasesHow to Evaluate AI Chats Using Conversation Coherence EvaluatorChain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous Sources
Hide in Main Feed
Ready to Publish
Nov 27, 2023
Jun 22, 2024 11:09 AM
Deficiency of Large Language Models in Finance: An Empirical Examination of HallucinationChain-of-Verification Reduces Hallucination in Large Language ModelsSelf-contradictory Hallucinations of Large Language Models: Evaluation, Detection and MitigationA Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language Models
Ready to Publish
May 25, 2023
Jun 22, 2024 11:08 AM
Semi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model ReasoningAI Safety: Necessary, but insufficient and possibly problematicDeficiency of Large Language Models in Finance: An Empirical Examination of HallucinationSiren's Song in the AI Ocean: A Survey on Hallucination in Large Language ModelsProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsA Comprehensive Survey on Instruction Following
Ready to Publish
Sep 3, 2023
Jun 22, 2024 11:08 AM
Semi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model ReasoningFine-tuning Language Models for FactualitySelf-contradictory Hallucinations of Large Language Models: Evaluation, Detection and MitigationA Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
Ready to Publish
Sep 30, 2023
Jun 22, 2024 11:07 AM
Search-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive TasksProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsSemi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model ReasoningActive Retrieval Augmented GenerationA Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language Models
Ready to Publish
Nov 9, 2023
Jun 22, 2024 11:06 AM
Siren's Song in the AI Ocean: A Survey on Hallucination in Large Language ModelsEntGPT: Linking Generative Large Language Models with Knowledge BasesPrompt Design and Engineering: Introduction and Advanced MethodsEnhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4Large Language Models as Analogical Reasoners
Ready to Publish
Sep 20, 2023
Jun 22, 2024 11:06 AM
One Small Step for Generative AI, One Giant Leap for AGI: A Complete Survey on ChatGPT in AIGC EraReasoning with Language Model Prompting: A SurveyTowards Reasoning in Large Language Models: A Survey
Ready to Publish
Apr 12, 2023
Jun 22, 2024 11:03 AM
Reprompting: Automated Chain-of-Thought Prompt Inference Through Gibbs SamplingFlatness-Aware Prompt Selection Improves Accuracy and Sample EfficiencySatLM: Satisfiability-Aided Language Models Using Declarative PromptingREFINER: Reasoning Feedback on Intermediate RepresentationsReflexion: Language Agents with Verbal Reinforcement LearningCAMEL: Communicative Agents for "Mind" Exploration of Large Language Model SocietySelf-Refine: Iterative Refinement with Self-FeedbackFairness-guided Few-shot Prompting for Large Language ModelsContext-faithful Prompting for Large Language ModelsStructure Pretraining and Prompt Tuning for Knowledge Graph TransferCoTEVer: Chain of Thought Prompting Annotation Toolkit for Explanation VerificationLarger language models do in-context learning differentlyOpenICL: An Open-Source Framework for In-context LearningDynamic Prompting: A Unified Framework for Prompt Tuning
Ready to Publish
Jun 20, 2024 01:42 PM
Ready to Publish
May 9, 2024
Featured
Jun 20, 2024 01:41 PM
Breaking Down the Defenses: A Comparative Survey of Attacks on Large Language ModelsUniversal and Transferable Adversarial Attacks on Aligned Language ModelsEver: Mitigating Hallucination in Large Language Models through Real-Time Verification and RectificationFrom Noise to Clarity: Unraveling the Adversarial Suffix of Large Language Model Attacks via Translation of Text EmbeddingsMistral 7B: Foundation Model Research Paper SummaryWizardLM: Empowering Large Language Models to Follow Complex InstructionsEntGPT: Linking Generative Large Language Models with Knowledge BasesCYBERSECEVAL 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language ModelsLanguage Prompt for Autonomous Driving
Ready to Publish
Feb 9, 2024
Jun 20, 2024 01:40 PM
From Noise to Clarity: Unraveling the Adversarial Suffix of Large Language Model Attacks via Translation of Text EmbeddingsUniversal and Transferable Adversarial Attacks on Aligned Language ModelsPrompt Injection: Different Attacks and Defensive TechniquesFactuality of Large Language Models in the Year 2024KnowGPT: Knowledge Injection for Large Language ModelsProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsA Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
Ready to Publish
Nov 13, 2023
Jun 20, 2024 01:40 PM
Many-Shot Jailbreaking (Anthropic Research)Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and RectificationDirect Preference Optimization: Your Language Model is Secretly a Reward ModelChain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous SourcesSelf-contradictory Hallucinations of Large Language Models: Evaluation, Detection and MitigationSiren's Song in the AI Ocean: A Survey on Hallucination in Large Language ModelsAutoHall: Automated Hallucination Dataset Generation for Large Language Models
Ready to Publish
May 22, 2023
Jun 20, 2024 01:39 PM
Factuality of Large Language Models in the Year 2024Chain-of-Knowledge: Grounding Large Language Models via Dynamic Knowledge Adapting over Heterogeneous SourcesSemi-Structured Chain-of-Thought: Integrating Multiple Sources of Knowledge for Improved Language Model ReasoningFine-tuning Language Models for Factuality
Ready to Publish
Apr 28, 2023
Jun 20, 2024 01:39 PM
WizardLM: Empowering Large Language Models to Follow Complex InstructionsEver: Mitigating Hallucination in Large Language Models through Real-Time Verification and RectificationProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsAutoHall: Automated Hallucination Dataset Generation for Large Language ModelsActive Retrieval Augmented Generation
Ready to Publish
Nov 23, 2023
Jun 20, 2024 01:39 PM
Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and MitigationSearch-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive TasksEntGPT: Linking Generative Large Language Models with Knowledge BasesAutoHall: Automated Hallucination Dataset Generation for Large Language ModelsA Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language ModelsPrompt Design and Engineering: Introduction and Advanced Methods
Ready to Publish
Jan 24, 2024
Jun 20, 2024 01:39 PM
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open QuestionsActive Retrieval Augmented GenerationProbabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex QuestionsLarge Language Models as Analogical ReasonersLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsRe-Reading Improves Reasoning in Large Language ModelsEnhancing Large Language Models Against Inductive Instructions with Dual-critique PromptingPost Hoc Explanations of Language Models Can Improve Language ModelsTree of Thoughts: Deliberate Problem Solving with Large Language ModelsUPRISE: Universal Prompt Retrieval for Improving Zero-Shot EvaluationModel-tuning Via Prompts Makes NLP Models Adversarially RobustInvestigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction FollowingAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled DataGlobal Prompt Cell: A Portable Control Module for Effective Prompt Tuning
Ready to Publish
Sep 23, 2023
Jun 20, 2024 01:38 PM
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open QuestionsA Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language ModelsActive Retrieval Augmented GenerationLarge Language Models as Analogical ReasonersLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsEnhancing Large Language Models Against Inductive Instructions with Dual-critique PromptingModel-tuning Via Prompts Makes NLP Models Adversarially RobustInvestigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction FollowingAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
Ready to Publish
Dec 26, 2023
Jun 20, 2024 01:38 PM
A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open QuestionsA Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language ModelsActive Retrieval Augmented GenerationLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersRe-Reading Improves Reasoning in Large Language Models
Ready to Publish
Oct 3, 2023
Jun 20, 2024 01:38 PM
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicA Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open QuestionsPrompt Design and Engineering: Introduction and Advanced Methods
Ready to Publish
Oct 9, 2023
Jun 20, 2024 01:38 PM
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicPrompt Design and Engineering: Introduction and Advanced MethodsPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4Connecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersRe-Reading Improves Reasoning in Large Language ModelsSkeleton-of-Thought: Prompting LLMs for Efficient Parallel GenerationEnhancing Large Language Models Against Inductive Instructions with Dual-critique PromptingUPRISE: Universal Prompt Retrieval for Improving Zero-Shot EvaluationModel-tuning Via Prompts Makes NLP Models Adversarially RobustInvestigating the Effectiveness of Task-Agnostic Prefix Prompt for Instruction Following
Ready to Publish
Sep 15, 2023
Jun 20, 2024 01:37 PM
LLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4Skeleton-of-Thought: Prompting LLMs for Efficient Parallel GenerationAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled Data
Ready to Publish
Sep 12, 2023
Jun 20, 2024 01:36 PM
LLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsPrincipled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4Prompt Design and Engineering: Introduction and Advanced MethodsSkeleton-of-Thought: Prompting LLMs for Efficient Parallel GenerationPost Hoc Explanations of Language Models Can Improve Language Models
Ready to Publish
Jul 28, 2023
Jun 20, 2024 01:36 PM
Re-Reading Improves Reasoning in Large Language ModelsConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsPost Hoc Explanations of Language Models Can Improve Language Models
Ready to Publish
May 23, 2023
Jun 20, 2024 01:36 PM
Prompt Design and Engineering: Introduction and Advanced MethodsLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsEnhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicTree of Thoughts: Deliberate Problem Solving with Large Language ModelsKnowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
Ready to Publish
May 19, 2023
Jun 20, 2024 01:33 PM
Re-Reading Improves Reasoning in Large Language ModelsSkeleton-of-Thought: Prompting LLMs for Efficient Parallel GenerationPrompt Design and Engineering: Introduction and Advanced MethodsTree of Thoughts: Deliberate Problem Solving with Large Language ModelsKnowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
Ready to Publish
May 17, 2023
Jun 20, 2024 01:33 PM
Prompt Design and Engineering: Introduction and Advanced MethodsPost Hoc Explanations of Language Models Can Improve Language ModelsEnhancing Large Language Models Against Inductive Instructions with Dual-critique PromptingKnowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
Ready to Publish
May 17, 2023
Jun 20, 2024 01:33 PM
Enhancing Large Language Models Against Inductive Instructions with Dual-critique PromptingPost Hoc Explanations of Language Models Can Improve Language ModelsTree of Thoughts: Deliberate Problem Solving with Large Language ModelsUPRISE: Universal Prompt Retrieval for Improving Zero-Shot Evaluation
Ready to Publish
Mar 18, 2023
Jun 20, 2024 01:32 PM
A Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language ModelsSelf-contradictory Hallucinations of Large Language Models: Evaluation, Detection and MitigationActive Retrieval Augmented GenerationGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningRevisiting Automated Prompting: Are We Actually Doing Better?NN Prompting: Beyond-Context Learning with Calibration-Free Nearest Neighbor InferenceVisual-Language Prompt Tuning with Knowledge-guided Context OptimizationEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited Data
Ready to Publish
Mar 15, 2023
Jun 20, 2024 01:32 PM
LLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsPrompt Design and Engineering: Introduction and Advanced MethodsKnowledge Card: Filling LLMs' Knowledge Gaps with Plug-in Specialized Language Models
Ready to Publish
Mar 13, 2023
Jun 20, 2024 01:31 PM
Prompt Design and Engineering: Introduction and Advanced MethodsLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsEnhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicExploring LLM-based Agents for Root Cause AnalysisReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and Beyond
Ready to Publish
Feb 28, 2023
Jun 20, 2024 01:31 PM
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicLLMLingua: Compressing Prompts for Accelerated Inference of Large Language ModelsPrompt Design and Engineering: Introduction and Advanced MethodsReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and Beyond
Ready to Publish
Feb 24, 2023
Jun 20, 2024 01:31 PM
Enhancing Zero-Shot Chain-of-Thought Reasoning in Large Language Models through LogicPrompt Design and Engineering: Introduction and Advanced MethodsConnecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersExploring LLM-based Agents for Root Cause AnalysisFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and Beyond
Ready to Publish
May 23, 2023
Jun 20, 2024 01:24 PM
Reinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyExploring LLM-based Agents for Root Cause AnalysisA Bibliometric Review of Large Language Models Research from 2017 to 2023From Sparse to Dense: GPT-4 Summarization with Chain of Density PromptingHierarchical Prompting Assists Large Language Model on Web NavigationInteractive Natural Language Processing
Ready to Publish
Apr 20, 2022
Jun 20, 2024 01:24 PM
Towards Reasoning in Large Language Models: A SurveyReasoning with Language Model Prompting: A SurveyEmergent Abilities of Large Language Models
Ready to Publish
Jul 28, 2021
Jun 20, 2024 01:24 PM
Ready to Publish
Sep 8, 2023
Jun 20, 2024 01:23 PM
Towards Reasoning in Large Language Models: A SurveyReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondJailbreaking ChatGPT via Prompt Engineering: An Empirical Study
Ready to Publish
Aug 18, 2023
Jun 20, 2024 01:23 PM
Reasoning with Language Model Prompting: A SurveyTowards Reasoning in Large Language Models: A SurveyGraph of Thoughts: Solving Elaborate Problems with Large Language ModelsLess Likely Brainstorming: Using Language Models to Generate Alternative HypothesesUniversality and Limitations of Prompt TuningMultiTool-CoT: GPT-3 Can Use Multiple External Tools with Chain of Thought PromptingReasoning with Language Model is Planning with World ModelBetter Zero-Shot Reasoning with Self-Adaptive PromptingLet's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning and Coding with LLMsTreePrompt: Learning to Compose Tree Prompts for Explainable Visual Grounding
Ready to Publish
May 31, 2023
Jun 20, 2024 01:22 PM
Towards Reasoning in Large Language Models: A SurveyReasoning with Language Model Prompting: A SurveyReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and Beyond
Ready to Publish
May 23, 2023
Jun 20, 2024 01:22 PM
Reinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondA Bibliometric Review of Large Language Models Research from 2017 to 2023Few-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
May 23, 2023
Jun 20, 2024 01:22 PM
Graph of Thoughts: Solving Elaborate Problems with Large Language ModelsReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
Ready to Publish
May 23, 2023
Jun 20, 2024 01:21 PM
Jailbreaking ChatGPT via Prompt Engineering: An Empirical StudyA Bibliometric Review of Large Language Models Research from 2017 to 2023Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and Beyond
Ready to Publish
Mar 31, 2023
Jun 20, 2024 01:12 PM
Boosted Prompt Ensembles for Large Language ModelsGlobal Prompt Cell: A Portable Control Module for Effective Prompt TuningWhy think step by step? Reasoning emerges from the locality of experience
Ready to Publish
Mar 6, 2023
Jun 20, 2024 01:12 PM
Boosted Prompt Ensembles for Large Language ModelsRevisiting Automated Prompting: Are We Actually Doing Better?CoTEVer: Chain of Thought Prompting Annotation Toolkit for Explanation VerificationART: Automatic multi-step reasoning and tool-use for large language modelsMultitask Prompt Tuning Enables Parameter-Efficient Transfer LearningAlphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Ready to Publish
Feb 28, 2023
Jun 20, 2024 01:11 PM
How Robust is GPT-3.5 to Predecessors? A Comprehensive Study on Language Understanding TasksPrompt, Generate, then Cache: Cascade of Foundation Models makes Strong Few-shot LearnersEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataActive Prompting with Chain-of-Thought for Large Language ModelsNot what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection
Ready to Publish
Feb 21, 2023
Jun 20, 2024 01:11 PM
Language Is Not All You Need: Aligning Perception with Language ModelsNot what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt InjectionActive Prompting with Chain-of-Thought for Large Language ModelsGuiding Large Language Models via Directional Stimulus PromptingHow Does In-Context Learning Help Prompt Tuning?A-la-carte Prompt Tuning (APT): Combining Distinct Data Via Composable Prompting
Ready to Publish
Oct 9, 2023
Jun 20, 2024 01:11 PM
A Prompt Pattern Catalog to Enhance Prompt Engineering with ChatGPTLanguage Is Not All You Need: Aligning Perception with Language ModelsActive Prompting with Chain-of-Thought for Large Language ModelsHow Does In-Context Learning Help Prompt Tuning?
Ready to Publish
Feb 18, 2023
Jun 20, 2024 01:10 PM
How Does In-Context Learning Help Prompt Tuning?Chain of Hindsight Aligns Language Models with FeedbackEffectiveness of Data Augmentation for Parameter Efficient Tuning with Limited DataThe Capacity for Moral Self-Correction in Large Language Models
Ready to Publish
Feb 25, 2023
Jun 20, 2024 01:06 PM
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt ConstraintsA-la-carte Prompt Tuning (APT): Combining Distinct Data Via Composable PromptingHow Does In-Context Learning Help Prompt Tuning?SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsEvaluating the Robustness of Discrete PromptsHard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoveryProgressive Prompts: Continual Learning for Language ModelsBatch Prompting: Efficient Inference with Large Language Model APIsDemonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLPOn Second Thought, Let's Not Think Step by Step! Bias and Toxicity in Zero-Shot ReasoningConstitutional AI: Harmlessness from AI FeedbackLarge Language Models are reasoners with Self-Verification
Ready to Publish
Feb 18, 2023
Jun 20, 2024 01:04 PM
Bounding the Capabilities of Large Language Models in Open Text Generation with Prompt ConstraintsScalable Prompt Generation for Semi-supervised Learning with Language ModelsHow Does In-Context Learning Help Prompt Tuning?SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsEvaluating the Robustness of Discrete PromptsCompositional Exemplars for In-context Learning
SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource Domains
Ready to Publish
Feb 14, 2023
Jun 20, 2024 01:04 PM
The Capacity for Moral Self-Correction in Large Language ModelsGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksA-la-carte Prompt Tuning (APT): Combining Distinct Data Via Composable PromptingCompositional Exemplars for In-context LearningMultimodal Chain-of-Thought Reasoning in Language ModelsLarge Language Models Can Be Easily Distracted by Irrelevant ContextSynthetic Prompting: Generating Chain-of-Thought Demonstrations for Large Language Models
Ready to Publish
Feb 11, 2023
Jun 20, 2024 01:04 PM
GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksThe Capacity for Moral Self-Correction in Large Language ModelsEvaluating the Robustness of Discrete PromptsCompositional Exemplars for In-context LearningMultimodal Chain-of-Thought Reasoning in Language ModelsDemonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP
Ready to Publish
Jun 20, 2023
Jun 20, 2024 01:03 PM
SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsThe Capacity for Moral Self-Correction in Large Language ModelsEvaluating the Robustness of Discrete PromptsHard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoveryMultimodal Chain-of-Thought Reasoning in Language ModelsRetrieval-Augmented Thought Process as Sequential Decision Making
Ready to Publish
Jun 1, 2023
Jun 20, 2024 12:55 PM
GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksCompositional Exemplars for In-context LearningHard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoveryLarge Language Models Can Be Easily Distracted by Irrelevant ContextBatch Prompting: Efficient Inference with Large Language Model APIsOn Second Thought, Let's Not Think Step by Step! Bias and Toxicity in Zero-Shot Reasoning
Ready to Publish
Feb 17, 2023
Jun 20, 2024 12:54 PM
SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsEvaluating the Robustness of Discrete PromptsCompositional Exemplars for In-context LearningLarge Language Models Can Be Easily Distracted by Irrelevant ContextSynthetic Prompting: Generating Chain-of-Thought Demonstrations for Large Language ModelsProgressive Prompts: Continual Learning for Language ModelsBatch Prompting: Efficient Inference with Large Language Model APIsRetrieval-Augmented Thought Process as Sequential Decision Making
Ready to Publish
Jun 6, 2023
Jun 20, 2024 12:54 PM
Multimodal Chain-of-Thought Reasoning in Language ModelsSwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsHard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoverySynthetic Prompting: Generating Chain-of-Thought Demonstrations for Large Language ModelsProgressive Prompts: Continual Learning for Language ModelsDemonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLPOn Second Thought, Let's Not Think Step by Step! Bias and Toxicity in Zero-Shot ReasoningUnleashing the potential of prompt engineering in Large Language Models: a comprehensive review
Ready to Publish
Feb 1, 2023
Jun 20, 2024 12:54 PM
SwitchPrompt: Learning Domain-Specific Gated Soft Prompts for Classification in Low-Resource DomainsMultimodal Chain-of-Thought Reasoning in Language ModelsLarge Language Models Can Be Easily Distracted by Irrelevant ContextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive review
Ready to Publish
Jan 29, 2023
Jun 20, 2024 12:54 PM
GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksMultimodal Chain-of-Thought Reasoning in Language ModelsLarge Language Models Can Be Easily Distracted by Irrelevant ContextThe Flan Collection: Designing Data and Methods for Effective Instruction Tuning
Ready to Publish
Oct 24, 2023
Jun 20, 2024 12:53 PM
Multimodal Chain-of-Thought Reasoning in Language ModelsHard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoveryGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksConstitutional AI: Harmlessness from AI FeedbackSuccessive Prompting for Decomposing Complex Questions
Ready to Publish
Jan 23, 2023
Jun 20, 2024 12:53 PM
GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksLarge Language Models Can Be Easily Distracted by Irrelevant ContextEvaluating the Robustness of Discrete PromptsSuccessive Prompting for Decomposing Complex QuestionsLarge Language Models are reasoners with Self-Verification
Ready to Publish
Jun 4, 2023
Jun 20, 2024 12:53 PM
Hard Prompts Made Easy: Gradient-Based Discrete Optimization for Prompt Tuning and DiscoveryLarge Language Models Can Be Easily Distracted by Irrelevant ContextGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksConstitutional AI: Harmlessness from AI Feedback
Ready to Publish
Dec 15, 2022
Jun 20, 2024 12:52 PM
GraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksOn Second Thought, Let's Not Think Step by Step! Bias and Toxicity in Zero-Shot ReasoningBatch Prompting: Efficient Inference with Large Language Model APIsSuccessive Prompting for Decomposing Complex QuestionsLarge Language Models are reasoners with Self-VerificationDemystifying Chains, Trees, and Graphs of Thoughts
Ready to Publish
Dec 8, 2022
Jun 20, 2024 12:51 PM
Batch Prompting: Efficient Inference with Large Language Model APIsConstitutional AI: Harmlessness from AI FeedbackDemonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLPThe Flan Collection: Designing Data and Methods for Effective Instruction Tuning
Ready to Publish
Dec 19, 2022
Jun 20, 2024 12:51 PM
Constitutional AI: Harmlessness from AI FeedbackGraphPrompt: Unifying Pre-Training and Downstream Tasks for Graph Neural NetworksDemonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLPDemystifying Chains, Trees, and Graphs of Thoughts
Ready to Publish
Feb 14, 2023
Jun 20, 2024 12:51 PM
The Flan Collection: Designing Data and Methods for Effective Instruction TuningSuccessive Prompting for Decomposing Complex QuestionsProgressive Prompts: Continual Learning for Language ModelsDemystifying Chains, Trees, and Graphs of ThoughtsAlgorithm of Thoughts: Enhancing Exploration of Ideas in Large Language Models
Ready to Publish
Apr 5, 2024
Jun 20, 2024 12:50 PM
The Flan Collection: Designing Data and Methods for Effective Instruction TuningLarge Language Models are reasoners with Self-VerificationConstitutional AI: Harmlessness from AI FeedbackAlgorithm of Thoughts: Enhancing Exploration of Ideas in Large Language Models
Ready to Publish
Oct 27, 2023
Jun 20, 2024 12:50 PM
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewLarge Language Models Can Be Easily Distracted by Irrelevant ContextSynthetic Prompting: Generating Chain-of-Thought Demonstrations for Large Language ModelsAlgorithm of Thoughts: Enhancing Exploration of Ideas in Large Language ModelsEmpowering Multi-step Reasoning across Languages via Tree-of-ThoughtsTree of Attacks: Jailbreaking Black-Box LLMs AutomaticallyLarge Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape SamplesGuReT: Distinguishing Guilt and Regret related TextRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalBoosting of Thoughts: Trial-and-Error Problem Solving with Large Language ModelsPathFinder: Guided Search over Multi-Step Reasoning PathsSPROUT: Authoring Programming Tutorials with Interactive Visualization of Large Language Model Generation ProcessNLPBench: Evaluating Large Language Models on Solving NLP ProblemsTree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question Answering
Ready to Publish
Sep 28, 2023
Jun 20, 2024 12:50 PM
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewDemystifying Chains, Trees, and Graphs of ThoughtsThe Flan Collection: Designing Data and Methods for Effective Instruction TuningEverything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationBoosting Logical Reasoning in Large Language Models through a New Framework: The Graph of ThoughtTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual Reasoning
Ready to Publish
Feb 12, 2024
Jun 20, 2024 12:49 PM
Retrieval-Augmented Thought Process as Sequential Decision MakingMultimodal Chain-of-Thought Reasoning in Language ModelsCompositional Exemplars for In-context LearningEverything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationBoosting Logical Reasoning in Large Language Models through a New Framework: The Graph of ThoughtTree of Attacks: Jailbreaking Black-Box LLMs Automatically
Ready to Publish
Feb 23, 2024
Jun 20, 2024 12:49 PM
Algorithm of Thoughts: Enhancing Exploration of Ideas in Large Language ModelsRetrieval-Augmented Thought Process as Sequential Decision MakingEverything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationEmpowering Multi-step Reasoning across Languages via Tree-of-ThoughtsBoosting Logical Reasoning in Large Language Models through a New Framework: The Graph of ThoughtTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual ReasoningLarge Language Model Guided Tree-of-ThoughtMACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical Problems
Ready to Publish
Feb 9, 2024
Jun 20, 2024 12:49 PM
Dynamic Prompting: A Unified Framework for Prompt TuningAlphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingLarger language models do in-context learning differentlyEmpowering Multi-step Reasoning across Languages via Tree-of-ThoughtsTree of Attacks: Jailbreaking Black-Box LLMs AutomaticallyTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual ReasoningLarge Language Model Guided Tree-of-ThoughtMACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical Problems
Ready to Publish
Apr 19, 2024
Jun 20, 2024 12:48 PM
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewAlphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingEverything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation
Ready to Publish
Aug 16, 2023
Jun 20, 2024 12:48 PM
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationRetrieval-Augmented Thought Process as Sequential Decision MakingAlgorithm of Thoughts: Enhancing Exploration of Ideas in Large Language ModelsFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Feb 21, 2024
Jun 20, 2024 12:48 PM
Alphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewRetrieval-Augmented Thought Process as Sequential Decision Making
Ready to Publish
Aug 21, 2023
Jun 20, 2024 12:48 PM
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationAlphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingAlgorithm of Thoughts: Enhancing Exploration of Ideas in Large Language ModelsLarge Language Model Guided Tree-of-ThoughtMACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical ProblemsGuReT: Distinguishing Guilt and Regret related Text
Ready to Publish
May 15, 2023
Jun 20, 2024 12:47 PM
Everything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationAlphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual ReasoningLarge Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape SamplesGuReT: Distinguishing Guilt and Regret related TextFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Apr 6, 2024
Jun 20, 2024 12:47 PM
Alphazero-like Tree-Search can Guide Large Language Model Decoding and TrainingEverything of Thoughts: Defying the Law of Penrose Triangle for Thought GenerationTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual ReasoningLarge Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape Samples
Ready to Publish
Feb 12, 2024
Jun 20, 2024 12:47 PM
MACM: Utilizing a Multi-Agent System for Condition Mining in Solving Complex Mathematical ProblemsUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewLarge Language Model Guided Tree-of-ThoughtRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalTree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question AnsweringEnhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice Guidelines
Ready to Publish
Jan 29, 2024
Jun 20, 2024 12:46 PM
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewLarge Language Model Guided Tree-of-ThoughtTree-of-Mixed-Thought: Combining Fast and Slow Thinking for Multi-hop Visual ReasoningFounder-GPT: Self-play to evaluate the Founder-Idea fitRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalBoosting of Thoughts: Trial-and-Error Problem Solving with Large Language ModelsAI Chain on Large Language Model for Unsupervised Control Flow Graph Generation for Statically-Typed Partial CodePathFinder: Guided Search over Multi-Step Reasoning PathsSPROUT: Authoring Programming Tutorials with Interactive Visualization of Large Language Model Generation ProcessNLPBench: Evaluating Large Language Models on Solving NLP ProblemsSelf-Taught Optimizer (STOP): Recursively Self-Improving Code GenerationKnowledge-Driven CoT: Exploring Faithful Reasoning in LLMs for Knowledge-intensive Question AnsweringTree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question AnsweringChain-of-Thought Reasoning is a Policy Improvement OperatorEnhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice GuidelinesRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsTemporal Data Meets LLM -- Explainable Financial Time Series ForecastingInferring Properties of Graph Neural Networks
Ready to Publish
Dec 20, 2023
Jun 20, 2024 12:45 PM
GuReT: Distinguishing Guilt and Regret related TextLarge Language Model Guided Tree-of-ThoughtBoosting Logical Reasoning in Large Language Models through a New Framework: The Graph of ThoughtBoosting of Thoughts: Trial-and-Error Problem Solving with Large Language ModelsAI Chain on Large Language Model for Unsupervised Control Flow Graph Generation for Statically-Typed Partial CodeSelf-Taught Optimizer (STOP): Recursively Self-Improving Code GenerationKnowledge-Driven CoT: Exploring Faithful Reasoning in LLMs for Knowledge-intensive Question AnsweringAutonomous Tree-search Ability of Large Language ModelsLLM Guided Evolution -- The Automation of Models Advancing ModelsRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsOn the Empirical Complexity of Reasoning and Planning in LLMsSTAMP: Differentiable Task and Motion Planning via Stein Variational Gradient DescentRoT: Enhancing Large Language Models with Reflection on Search TreesDiffusionGPT: LLM-Driven Text-to-Image Generation System
Ready to Publish
May 10, 2024
Jun 20, 2024 12:44 PM
Large Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape SamplesGuReT: Distinguishing Guilt and Regret related TextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewChain-of-Thought Reasoning is a Policy Improvement OperatorLLM Guided Evolution -- The Automation of Models Advancing ModelsGTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic EvaluationsAnalyzing Toxicity in Deep Conversations: A Reddit Case StudyRoT: Enhancing Large Language Models with Reflection on Search TreesPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationText2MDT: Extracting Medical Decision Trees from Medical Texts
Ready to Publish
Feb 17, 2024
Jun 20, 2024 12:44 PM
GuReT: Distinguishing Guilt and Regret related TextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewFounder-GPT: Self-play to evaluate the Founder-Idea fitAI Chain on Large Language Model for Unsupervised Control Flow Graph Generation for Statically-Typed Partial CodePathFinder: Guided Search over Multi-Step Reasoning Paths
Ready to Publish
Jun 1, 2023
Jun 20, 2024 12:44 PM
GuReT: Distinguishing Guilt and Regret related TextFounder-GPT: Self-play to evaluate the Founder-Idea fitBoosting of Thoughts: Trial-and-Error Problem Solving with Large Language Models
Ready to Publish
Dec 12, 2023
Jun 20, 2024 12:43 PM
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewGuReT: Distinguishing Guilt and Regret related TextBoosting of Thoughts: Trial-and-Error Problem Solving with Large Language ModelsSPROUT: Authoring Programming Tutorials with Interactive Visualization of Large Language Model Generation ProcessNLPBench: Evaluating Large Language Models on Solving NLP ProblemsSelf-Taught Optimizer (STOP): Recursively Self-Improving Code GenerationKnowledge-Driven CoT: Exploring Faithful Reasoning in LLMs for Knowledge-intensive Question AnsweringAutonomous Tree-search Ability of Large Language ModelsEnhancing Large Language Models for Clinical Decision Support by Incorporating Clinical Practice GuidelinesEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningOn the Empirical Complexity of Reasoning and Planning in LLMsGTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic EvaluationsAnalyzing Toxicity in Deep Conversations: A Reddit Case StudyRoT: Enhancing Large Language Models with Reflection on Search TreesInferring Properties of Graph Neural NetworksLayoutLLM: Layout Instruction Tuning with Large Language Models for Document UnderstandingPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationText2MDT: Extracting Medical Decision Trees from Medical Texts
Ready to Publish
Dec 4, 2023
Jun 20, 2024 12:43 PM
GuReT: Distinguishing Guilt and Regret related TextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewPathFinder: Guided Search over Multi-Step Reasoning PathsAutonomous Tree-search Ability of Large Language Models
Ready to Publish
Oct 19, 2023
Jun 20, 2024 12:42 PM
GuReT: Distinguishing Guilt and Regret related TextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewPathFinder: Guided Search over Multi-Step Reasoning Paths
Ready to Publish
Mar 1, 2024
Jun 20, 2024 12:42 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsGuReT: Distinguishing Guilt and Regret related TextFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Oct 28, 2023
Jun 20, 2024 12:42 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsGuReT: Distinguishing Guilt and Regret related TextFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Oct 14, 2023
Jun 20, 2024 12:42 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsSPROUT: Authoring Programming Tutorials with Interactive Visualization of Large Language Model Generation ProcessFounder-GPT: Self-play to evaluate the Founder-Idea fit
Tree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question Answering
Ready to Publish
Apr 22, 2024
Jun 20, 2024 12:41 PM
Large Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape SamplesGuReT: Distinguishing Guilt and Regret related TextUnleashing the potential of prompt engineering in Large Language Models: a comprehensive reviewChain-of-Thought Reasoning is a Policy Improvement OperatorLLM Guided Evolution -- The Automation of Models Advancing Models
Ready to Publish
Nov 8, 2023
Jun 20, 2024 12:41 PM
Tree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question AnsweringGuReT: Distinguishing Guilt and Regret related TextRNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
Ready to Publish
Mar 18, 2024
Jun 20, 2024 12:41 PM
Founder-GPT: Self-play to evaluate the Founder-Idea fitRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalTree of Reviews: A Tree-based Dynamic Iterative Retrieval Framework for Multi-hop Question Answering
Ready to Publish
Jan 23, 2024
Jun 20, 2024 12:41 PM
GuReT: Distinguishing Guilt and Regret related TextPathFinder: Guided Search over Multi-Step Reasoning PathsLarge Language Models are Few-shot Generators: Proposing Hybrid Prompt Algorithm To Generate Webshell Escape Samples
Ready to Publish
Apr 18, 2024
Jun 20, 2024 12:40 PM
Founder-GPT: Self-play to evaluate the Founder-Idea fitGuReT: Distinguishing Guilt and Regret related TextRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningTemporal Data Meets LLM -- Explainable Financial Time Series ForecastingSTAMP: Differentiable Task and Motion Planning via Stein Variational Gradient DescentInferring Properties of Graph Neural NetworksLayoutLLM: Layout Instruction Tuning with Large Language Models for Document UnderstandingDiffusionGPT: LLM-Driven Text-to-Image Generation SystemWho Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human PreferencesResearDesign Guidelines for Prompt Engineering Text-to-Image Generative ModelsChain-of-Thought Prompting Elicits Reasoning in Large Language ModelsSelf-Consistency Improves Chain of Thought Reasoning in Language ModelsLarge Language Models are Zero-Shot ReasonersDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical Reasoning
Ready to Publish
Jan 11, 2024
Jun 20, 2024 12:40 PM
RAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningPathFinder: Guided Search over Multi-Step Reasoning PathsOn the Empirical Complexity of Reasoning and Planning in LLMsGTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic EvaluationsTemporal Data Meets LLM -- Explainable Financial Time Series ForecastingSTAMP: Differentiable Task and Motion Planning via Stein Variational Gradient DescentAnalyzing Toxicity in Deep Conversations: A Reddit Case StudyLayoutLLM: Layout Instruction Tuning with Large Language Models for Document UnderstandingDiffusionGPT: LLM-Driven Text-to-Image Generation SystemPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationText2MDT: Extracting Medical Decision Trees from Medical TextsAutomatic Root Cause Analysis via Large Language Models for Cloud Incidents
Ready to Publish
Apr 17, 2024
Jun 20, 2024 12:40 PM
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningPathFinder: Guided Search over Multi-Step Reasoning PathsFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Feb 19, 2024
Jun 20, 2024 12:40 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningRNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
Ready to Publish
Jun 19, 2023
Jun 20, 2024 12:39 PM
GuReT: Distinguishing Guilt and Regret related TextRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented Reasoning
Ready to Publish
Jan 7, 2024
Jun 20, 2024 12:39 PM
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsFounder-GPT: Self-play to evaluate the Founder-Idea fit
Ready to Publish
Apr 11, 2024
Jun 20, 2024 12:39 PM
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningPathFinder: Guided Search over Multi-Step Reasoning PathsRNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
Ready to Publish
Apr 11, 2024
Jun 20, 2024 12:39 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsFounder-GPT: Self-play to evaluate the Founder-Idea fitRNNs are not Transformers (Yet): The Key Bottleneck on In-context Retrieval
Ready to Publish
Mar 2, 2024
Jun 20, 2024 12:38 PM
GuReT: Distinguishing Guilt and Regret related TextRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsPathFinder: Guided Search over Multi-Step Reasoning PathsSelf-Consistency Improves Chain of Thought Reasoning in Language ModelsPrompting GPT-3 To Be Reliable
Ready to Publish
Apr 8, 2024
Jun 20, 2024 12:38 PM
PathFinder: Guided Search over Multi-Step Reasoning PathsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsLarge Language Models are Zero-Shot Reasoners
Ready to Publish
Jan 18, 2024
Jun 20, 2024 12:38 PM
RAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningFounder-GPT: Self-play to evaluate the Founder-Idea fitResearDesign Guidelines for Prompt Engineering Text-to-Image Generative ModelsChain-of-Thought Prompting Elicits Reasoning in Large Language ModelsMaking Large Language Models Better Reasoners with Step-Aware VerifierDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical Reasoning
Ready to Publish
Dec 7, 2023
Jun 20, 2024 12:37 PM
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningPathFinder: Guided Search over Multi-Step Reasoning PathsRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalAutomatic Root Cause Analysis via Large Language Models for Cloud IncidentsWho Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human PreferencesResearDesign Guidelines for Prompt Engineering Text-to-Image Generative ModelsSelf-Consistency Improves Chain of Thought Reasoning in Language ModelsLarge Language Models are Zero-Shot ReasonersMaking Large Language Models Better Reasoners with Step-Aware Verifier
Ready to Publish
Jan 4, 2024
Jun 20, 2024 12:37 PM
Evidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningPathFinder: Guided Search over Multi-Step Reasoning PathsRNNs are not Transformers (Yet): The Key Bottleneck on In-context RetrievalAutomatic Root Cause Analysis via Large Language Models for Cloud IncidentsWho Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human PreferencesChain-of-Thought Prompting Elicits Reasoning in Large Language ModelsMaking Large Language Models Better Reasoners with Step-Aware VerifierDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical Reasoning
Ready to Publish
Nov 13, 2023
Jun 20, 2024 12:37 PM
PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationEvidence to Generate (E2G): A Single-agent Two-step Prompting for Context Grounded and Retrieval Augmented ReasoningText2MDT: Extracting Medical Decision Trees from Medical Texts
Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs with Human Preferences
Ready to Publish
Apr 18, 2024
Jun 20, 2024 12:37 PM
Text2MDT: Extracting Medical Decision Trees from Medical TextsRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization
Ready to Publish
Sep 28, 2023
Jun 20, 2024 12:23 PM
DiffusionGPT: LLM-Driven Text-to-Image Generation SystemRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt Optimization
Ready to Publish
Jan 10, 2023
Jun 20, 2024 12:23 PM
Text2MDT: Extracting Medical Decision Trees from Medical TextsRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsDiffusionGPT: LLM-Driven Text-to-Image Generation System
Ready to Publish
Mar 7, 2023
Jun 20, 2024 12:23 PM
Inferring Properties of Graph Neural NetworksRAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationPAL: Program-aided Language Models
Ready to Publish
Jan 29, 2023
Jun 20, 2024 12:22 PM
RAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsPromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationLayoutLLM: Layout Instruction Tuning with Large Language Models for Document Understanding
Ready to Publish
May 24, 2023
Jun 20, 2024 12:22 PM
PromptAgent: Strategic Planning with Language Models Enables Expert-level Prompt OptimizationText2MDT: Extracting Medical Decision Trees from Medical TextsDiffusionGPT: LLM-Driven Text-to-Image Generation SystemDocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language Models
Ready to Publish
Mar 2, 2023
Jun 20, 2024 12:22 PM
RAGAR, Your Falsehood RADAR: RAG-Augmented Reasoning for Political Fact-Checking using Multimodal Large Language ModelsText2MDT: Extracting Medical Decision Trees from Medical TextsDiffusionGPT: LLM-Driven Text-to-Image Generation SystemDocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsLarge Language Models Are Human-Level Prompt EngineersMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
Feb 18, 2023
Jun 20, 2024 12:21 PM
Dynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical ReasoningDocPrompting: Generating Code by Retrieving the DocsMaking Large Language Models Better Reasoners with Step-Aware VerifierLarge Language Models Are Human-Level Prompt EngineersRecitation-Augmented Language ModelsPrompting GPT-3 To Be ReliableLanguage Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-ThoughtPrompt Engineering for Healthcare: Methodologies and ApplicationsPrompt Engineering a Prompt EngineerPrompt Engineering or Fine Tuning: An Empirical Assessment of Large Language Models in Automated Software Engineering TasksA Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt EngineeringPrompting AI Art: An Investigation into the Creative Skill of Prompt EngineeringA Systematic Survey of Prompt Engineering on Vision-Language Foundation ModelsUnderstanding prompt engineering may not require rethinking generalizationTo be or not to be? an exploration of continuously controllable prompt engineeringPEACE: Prompt Engineering Automation for CLIPSeg Enhancement in Aerial RoboticsPrompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesPrompt-Engineering and Transformer-based Question Generation and EvaluationPrompt Engineering-assisted Malware Dynamic Analysis Using GPT-4Cases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPTLarge Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation
Ready to Publish
Jan 27, 2023
Jun 20, 2024 12:21 PM
Making Large Language Models Better Reasoners with Step-Aware VerifierDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical ReasoningSelf-Consistency Improves Chain of Thought Reasoning in Language ModelsLarge Language Models Are Human-Level Prompt EngineersRecitation-Augmented Language ModelsDecomposed Prompting: A Modular Approach for Solving Complex TasksLanguage Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-ThoughtPrompt Engineering a Prompt EngineerBatch Calibration: Rethinking Calibration for In-Context Learning and Prompt EngineeringA Systematic Survey of Prompt Engineering on Vision-Language Foundation ModelsPEACE: Prompt Engineering Automation for CLIPSeg Enhancement in Aerial RoboticsPrompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesPrompt Engineering-assisted Malware Dynamic Analysis Using GPT-4Enhancing Medical Task Performance in GPT-4V: A Comprehensive Study on Prompt Engineering Strategies
Ready to Publish
Mar 10, 2023
Jun 20, 2024 12:20 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical ReasoningMachine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsReAct: Synergizing Reasoning and Acting in Language ModelsLanguage Models Are Greedy Reasoners: A Systematic Formal Analysis of Chain-of-ThoughtPrompt Engineering a Prompt EngineerPrompt Engineering or Fine Tuning: An Empirical Assessment of Large Language Models in Automated Software Engineering TasksBatch Calibration: Rethinking Calibration for In-Context Learning and Prompt EngineeringA Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt EngineeringPrompting AI Art: An Investigation into the Creative Skill of Prompt EngineeringPrompt Engineering-assisted Malware Dynamic Analysis Using GPT-4
Ready to Publish
May 8, 2023
Jun 20, 2024 12:20 PM
Large Language Models Are Human-Level Prompt EngineersMachine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsDynamic Prompt Learning via Policy Gradient for Semi-structured Mathematical ReasoningRecitation-Augmented Language ModelsReAct: Synergizing Reasoning and Acting in Language ModelsPrompting GPT-3 To Be ReliableDecomposed Prompting: A Modular Approach for Solving Complex TasksPrompt Engineering for Healthcare: Methodologies and ApplicationsBatch Calibration: Rethinking Calibration for In-Context Learning and Prompt EngineeringA Systematic Survey of Prompt Engineering on Vision-Language Foundation ModelsPEACE: Prompt Engineering Automation for CLIPSeg Enhancement in Aerial Robotics
Ready to Publish
Feb 16, 2023
Jun 20, 2024 12:20 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
Mar 10, 2023
Jun 20, 2024 12:19 PM
Large Language Models Are Human-Level Prompt EngineersMachine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsReAct: Synergizing Reasoning and Acting in Language ModelsPrompt Engineering for Healthcare: Methodologies and ApplicationsUnderstanding prompt engineering may not require rethinking generalizationPrompt-Engineering and Transformer-based Question Generation and EvaluationCases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPT
Ready to Publish
Feb 15, 2023
Jun 20, 2024 12:19 PM
DocPrompting: Generating Code by Retrieving the DocsInferring Properties of Graph Neural NetworksMachine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsDecomposed Prompting: A Modular Approach for Solving Complex TasksPrompt Engineering or Fine Tuning: An Empirical Assessment of Large Language Models in Automated Software Engineering TasksPrompting AI Art: An Investigation into the Creative Skill of Prompt EngineeringUnderstanding prompt engineering may not require rethinking generalizationTo be or not to be? an exploration of continuously controllable prompt engineeringPrompt-Engineering and Transformer-based Question Generation and EvaluationCases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPTEnhancing Medical Task Performance in GPT-4V: A Comprehensive Study on Prompt Engineering StrategiesLarge Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation
Ready to Publish
Apr 11, 2023
Jun 20, 2024 12:19 PM
Machine Generated Text: A Comprehensive Survey of Threat Models and Detection MethodsPrompting GPT-3 To Be ReliablePAL: Program-aided Language Models
Ready to Publish
Jan 26, 2023
Featured
Jun 20, 2024 12:18 PM
PAL: Program-aided Language ModelsDocPrompting: Generating Code by Retrieving the DocsLarge Language Models Are Human-Level Prompt Engineers
Ready to Publish
Mar 23, 2024
Jun 20, 2024 12:17 PM
ReAct: Synergizing Reasoning and Acting in Language ModelsDocPrompting: Generating Code by Retrieving the DocsMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
Feb 19, 2024
Jun 20, 2024 12:17 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsLarge Language Models Are Human-Level Prompt Engineers
Ready to Publish
Oct 11, 2023
Jun 20, 2024 12:17 PM
Prompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the DocsLarge Language Models Are Human-Level Prompt Engineers
Ready to Publish
Jan 24, 2024
Jun 20, 2024 12:16 PM
Large Language Models Are Human-Level Prompt EngineersPAL: Program-aided Language ModelsMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
Jul 3, 2023
Jun 20, 2024 12:16 PM
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt EngineeringDocPrompting: Generating Code by Retrieving the DocsLarge Language Models Are Human-Level Prompt EngineersTo be or not to be? an exploration of continuously controllable prompt engineeringLarge Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document Summarisation
Ready to Publish
Dec 3, 2023
Jun 20, 2024 12:16 PM
Prompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the DocsLarge Language Models Are Human-Level Prompt Engineers
Ready to Publish
Jul 24, 2023
Jun 20, 2024 12:15 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
Oct 6, 2023
Jun 20, 2024 12:14 PM
ReAct: Synergizing Reasoning and Acting in Language ModelsPrompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the Docs
Ready to Publish
Nov 16, 2023
Jun 20, 2024 12:14 PM
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt EngineeringPrompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the Docs
Ready to Publish
Dec 8, 2023
Jun 20, 2024 12:14 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsMachine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods
Ready to Publish
May 17, 2023
Jun 20, 2024 12:13 PM
Prompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesPAL: Program-aided Language ModelsDocPrompting: Generating Code by Retrieving the DocsEnhancing Medical Task Performance in GPT-4V: A Comprehensive Study on Prompt Engineering StrategiesSAMAug: Point Prompt Augmentation for Segment Anything ModelSAM on Medical Images: A Comprehensive Study on Three Prompt Modes
Ready to Publish
Oct 29, 2023
Jun 20, 2024 12:12 PM
ReAct: Synergizing Reasoning and Acting in Language ModelsPrompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the Docs
Ready to Publish
Dec 13, 2023
Jun 20, 2024 12:12 PM
DocPrompting: Generating Code by Retrieving the DocsPAL: Program-aided Language ModelsLarge Language Models Are Human-Level Prompt Engineers
Cases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPT
Ready to Publish
Jun 19, 2023
Jun 20, 2024 12:11 PM
ReAct: Synergizing Reasoning and Acting in Language ModelsPrompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the DocsExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Enhancing Medical Task Performance in GPT-4V: A Comprehensive Study on Prompt Engineering Strategies
Ready to Publish
Dec 12, 2023
Jun 20, 2024 12:10 PM
Prompting GPT-3 To Be ReliablePrompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesPAL: Program-aided Language ModelsMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
Nov 9, 2023
Jun 20, 2024 12:10 PM
A Survey on Segment Anything Model (SAM): Vision Foundation Model Meets Prompt EngineeringPrompting GPT-3 To Be ReliableDocPrompting: Generating Code by Retrieving the DocsExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification
Ready to Publish
Aug 14, 2023
Jun 20, 2024 12:09 PM
Large Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document SummarisationExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringCases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPTChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationAutomated Black-box Prompt Engineering for Personalized Text-to-Image GenerationChit-Chat or Deep Talk: Prompt Engineering for Process MiningSAMAug: Point Prompt Augmentation for Segment Anything ModelPrompt-Free Diffusion: Taking "Text" out of Text-to-Image Diffusion ModelsImproving ChatGPT Prompt for Code GenerationTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and PotentialSGL-PT: A Strong Graph Learner with Graph Prompt TuningState of What Art? A Call for Multi-Prompt LLM Evaluation
Ready to Publish
Jun 6, 2024
Jun 20, 2024 12:09 PM
Large Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document SummarisationEnhancing Medical Task Performance in GPT-4V: A Comprehensive Study on Prompt Engineering StrategiesCases of EFL Secondary Students' Prompt Engineering Pathways to Complete a Writing Task with ChatGPTChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classificationAutomated Black-box Prompt Engineering for Personalized Text-to-Image GenerationWordflow: Social Prompt Engineering for Large Language ModelsA Systematic Survey of Prompt Engineering in Large Language Models: Techniques and ApplicationsExploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveA Novel Approach for Rapid Development Based on ChatGPT and Prompt EngineeringChit-Chat or Deep Talk: Prompt Engineering for Process MiningSAMAug: Point Prompt Augmentation for Segment Anything ModelSAM on Medical Images: A Comprehensive Study on Three Prompt ModesPrompt-Free Diffusion: Taking "Text" out of Text-to-Image Diffusion ModelsDr ChatGPT, tell me what I want to hear: How prompt knowledge impacts health answer correctnessTowards Large-scale 3D Representation Learning with Multi-dataset Point Prompt TrainingPrompt Cache: Modular Attention Reuse for Low-Latency Inference
Ready to Publish
Mar 5, 2024
Jun 20, 2024 12:09 PM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringLarge Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document SummarisationLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classificationExploring Prompt Engineering Practices in the EnterpriseAutomated Black-box Prompt Engineering for Personalized Text-to-Image GenerationA Systematic Survey of Prompt Engineering in Large Language Models: Techniques and ApplicationsExploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveA Novel Approach for Rapid Development Based on ChatGPT and Prompt EngineeringChit-Chat or Deep Talk: Prompt Engineering for Process MiningImproving ChatGPT Prompt for Code Generation
Ready to Publish
Mar 23, 2024
Jun 20, 2024 12:08 PM
MedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationLarge Language Models and Prompt Engineering for Biomedical Query Focused Multi-Document SummarisationExploring Prompt Engineering Practices in the EnterpriseWordflow: Social Prompt Engineering for Large Language ModelsA Systematic Survey of Prompt Engineering in Large Language Models: Techniques and ApplicationsExploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveA Novel Approach for Rapid Development Based on ChatGPT and Prompt Engineering
Ready to Publish
Mar 13, 2024
Jun 20, 2024 12:08 PM
LAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classificationChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationExploring Prompt Engineering Practices in the EnterpriseWordflow: Social Prompt Engineering for Large Language Models
Ready to Publish
Mar 28, 2024
Jun 20, 2024 12:08 PM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
Ready to Publish
Jan 25, 2024
Jun 20, 2024 12:07 PM
MedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classificationExploring Prompt Engineering Practices in the Enterprise
Ready to Publish
Feb 5, 2024
Jun 20, 2024 12:07 PM
MedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification
Exploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspective
Ready to Publish
Feb 10, 2024
Jun 20, 2024 12:07 PM
LAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classificationChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsA study on Prompt Design, Advantages and Limitations of ChatGPT for Deep Learning Program RepairGraph-ToolFormer: To Empower LLMs with Graph Reasoning Ability via Prompt Augmented by ChatGPT
Ready to Publish
Dec 21, 2023
Jun 20, 2024 12:06 PM
MedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationLAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification
Ready to Publish
Jul 19, 2023
Jun 20, 2024 11:59 AM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineeringChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level GenerationHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsDr ChatGPT, tell me what I want to hear: How prompt knowledge impacts health answer correctnessTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and PotentialA study on Prompt Design, Advantages and Limitations of ChatGPT for Deep Learning Program RepairGraph-ToolFormer: To Empower LLMs with Graph Reasoning Ability via Prompt Augmented by ChatGPTState of What Art? A Call for Multi-Prompt LLM Evaluation
Ready to Publish
Mar 19, 2024
Jun 20, 2024 11:59 AM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringPrompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
Nov 27, 2023
Jun 20, 2024 11:50 AM
Chit-Chat or Deep Talk: Prompt Engineering for Process MiningHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsExploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveSAM on Medical Images: A Comprehensive Study on Three Prompt ModesPrompt-Free Diffusion: Taking "Text" out of Text-to-Image Diffusion ModelsImproving ChatGPT Prompt for Code GenerationDr ChatGPT, tell me what I want to hear: How prompt knowledge impacts health answer correctnessTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and PotentialSGL-PT: A Strong Graph Learner with Graph Prompt TuningTowards Large-scale 3D Representation Learning with Multi-dataset Point Prompt TrainingPrompt Cache: Modular Attention Reuse for Low-Latency InferenceState of What Art? A Call for Multi-Prompt LLM Evaluation
Ready to Publish
Apr 28, 2023
Jun 20, 2024 11:50 AM
How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsPrompt Engineering for Transformer-based Chemical Similarity Search Identifies Structurally Distinct Functional AnaloguesMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
Jun 1, 2023
Featured
Jun 20, 2024 11:49 AM
How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
May 15, 2023
Jun 20, 2024 11:49 AM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
Ready to Publish
Feb 23, 2023
Jun 20, 2024 11:48 AM
Chit-Chat or Deep Talk: Prompt Engineering for Process MiningHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
Mar 29, 2023
Jun 20, 2024 11:48 AM
Chit-Chat or Deep Talk: Prompt Engineering for Process MiningHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringA study on Prompt Design, Advantages and Limitations of ChatGPT for Deep Learning Program RepairGraph-ToolFormer: To Empower LLMs with Graph Reasoning Ability via Prompt Augmented by ChatGPTSGL-PT: A Strong Graph Learner with Graph Prompt TuningTowards Large-scale 3D Representation Learning with Multi-dataset Point Prompt TrainingPrompt Cache: Modular Attention Reuse for Low-Latency Inference
Ready to Publish
Apr 17, 2023
Jun 20, 2024 11:47 AM
Exploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveChit-Chat or Deep Talk: Prompt Engineering for Process MiningTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and Potential
Ready to Publish
May 11, 2023
Jun 20, 2024 11:47 AM
Exploring EFL students' prompt engineering in human-AI story writing: an Activity Theory perspectiveChit-Chat or Deep Talk: Prompt Engineering for Process MiningTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and Potential
Ready to Publish
Aug 15, 2023
Jun 20, 2024 11:46 AM
Exploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt EngineeringHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and Potential
Ready to Publish
Aug 18, 2023
Jun 20, 2024 11:43 AM
How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsTranslating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and PotentialMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
Apr 25, 2024
Jun 20, 2024 11:43 AM
Translating Radiology Reports into Plain Language using ChatGPT and GPT-4 with Prompt Learning: Promising Results, Limitations, and PotentialHow to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsMedPromptExtract (Medical Data Extraction Tool): Anonymization and Hi-fidelity Automated data extraction using NLP and prompt engineering
Ready to Publish
May 6, 2024
Jun 20, 2024 11:42 AM
How to Prompt LLMs for Text-to-SQL: A Study in Zero-shot, Single-domain, and Cross-domain SettingsChit-Chat or Deep Talk: Prompt Engineering for Process MiningExploring the Intersection of Large Language Models and Agent-Based Modeling via Prompt Engineering
Ready to Publish
Jun 11, 2024
Featured
Jun 11, 2024 11:05 AM
Ready to Publish
May 4, 2024
May 6, 2024 09:02 AM
Why think step by step? Reasoning emerges from the locality of experienceIntroducing Athina Prompt Management: A Powerful and Flexible Prompt Playground and CMS
Ready to Publish
Apr 17, 2024
May 6, 2024 08:50 AM
Factuality of Large Language Models in the Year 2024Fine-tuning Language Models for FactualityCookbook: How to set up Langchain tracing on Athina in 2 minutes
Ready to Publish
Mar 7, 2024
May 5, 2024 06:31 PM
Exploring LLM-based Agents for Root Cause AnalysisAutomatic Prompt Augmentation and Selection with Chain-of-Thought from Labeled DataModel-tuning Via Prompts Makes NLP Models Adversarially RobustReinforcement Learning in the Era of LLMs: What is Essential? What is needed? An RL Perspective on RLHF, Prompting, and BeyondFew-shot Fine-tuning vs. In-context Learning: A Fair Comparison and EvaluationJailbreaking ChatGPT via Prompt Engineering: An Empirical StudyHarnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondGlobal Prompt Cell: A Portable Control Module for Effective Prompt Tuning
Ready to Publish
Apr 22, 2024
Featured
Apr 23, 2024 02:07 AM
How to Evaluate AI Chats Using Conversation Coherence EvaluatorHow to evaluate your Llama Index query engine using Ragas evals + Athina AIHow to Use a Custom Grading Criteria to Evaluate LLM Responses (LLM-as-a-Judge)Introducing Athina Prompt Management: A Powerful and Flexible Prompt Playground and CMS
Ready to Publish
Apr 17, 2024
Apr 23, 2024 01:09 AM
CYBERSECEVAL 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language ModelsCookbook: How to set up Langchain tracing on Athina in 2 minutes
Ready to Publish
Apr 16, 2024
Featured
Apr 16, 2024 11:43 PM
Ready to Publish
Mar 30, 2024
iteralign
Apr 6, 2024 02:30 PM
Post-Semantic-Thinking: A Robust Strategy to Distill Reasoning Capacity from Large Language ModelsDirect Preference Optimization: Your Language Model is Secretly a Reward Model
Summary of Research Paper: IterAlign: Iterative Constitutional Alignment of Large Language Models
Summary of Research Paper: IterAlign: Iterative Constitutional Alignment of Large Language Models
Ready to Publish
Oct 10, 2023
Apr 4, 2024 04:51 AM
Ready to Publish
Feb 15, 2024
Apr 2, 2024 01:22 AM
Ready to Publish
Mar 13, 2024
Featured
If you're using Llama Index to work with advanced retrieval strategies, you're going to need a great evaluation setup.
Here's how you can use Athina's SDK to run Ragas evals on your Llama Index RAG pipeline.
Apr 2, 2024 01:21 AM
Post-Semantic-Thinking: A Robust Strategy to Distill Reasoning Capacity from Large Language ModelsCookbook: How to set up Langchain tracing on Athina in 2 minutes
Ready to Publish
Feb 6, 2024
Featured
Apr 2, 2024 01:21 AM
Ready to Publish
Mar 29, 2024
prompt-engineering-techniques
Featured
Prompt Engineering is an emerging field and the techniques are evolving every day. This article breaks down prompt engineering into a list of techniques.
Apr 2, 2024 01:21 AM
Ready to Publish
Mar 5, 2024
Featured
Apr 2, 2024 01:07 AM
Ready to Publish
Nov 29, 2023
Apr 2, 2024 01:06 AM
Ready to Publish
Oct 30, 2023
Apr 2, 2024 01:06 AM
Ready to Publish
Sep 20, 2023
Featured
Apr 2, 2024 12:15 AM
Ready to Publish
Mar 14, 2024
Apr 2, 2024 12:01 AM
Ready to Publish
Mar 31, 2024
Apr 2, 2024 12:01 AM