Kamalloo, E., Dziri, N., Clarke, C., & Rafiei, D. (2023). Evaluating Open-Domain Question Answering in the Era of Large Language Models ArXiv, abs/2305.06984. https://doi.org/10.48550/arXiv.2305.06984
References
Filter by:
2023
Conia, S., Li, M., Lee, D., Minhas, U. F., Ilyas, I., & Li, Y. (2023). Increasing Coverage and Precision of Textual Information in Multilingual Knowledge Graphs ArXiv, abs/2311.15781. https://doi.org/10.48550/ARXIV.2311.15781
Oladipo, A., Adeyemi, M., Ahia, O., Owodunni, A. T., Ogundepo, O., Adelani, D. I., & Lin, J. (2023). Better Quality Pre-Training Data and T5 Models for African Languages Presented at the Better Quality Pre-Training Data and T5 Models for African Languages conference. Retrieved from https://aclanthology.org/2023.emnlp-main.11
Rorseth, J., Godfrey, P., Golab, L., Kargar, M., Srivastava, D., & Szlichta, J. (2023). CREDENCE: Counterfactual Explanations for Document Ranking ArXiv, abs/2302.04983. https://doi.org/10.48550/arXiv.2302.04983
Buchanan, G. R., McKay, D., & Clarke, C. (2023). Made to Measure: A Workshop on Human-Centred Metrics for Information Seeking Presented at the Made to Measure: A Workshop on Human-Centred Metrics for Information Seeking Primary Tabs View conference. https://doi.org/10.1145/3576840.3578301
Kamalloo, E., Dziri, N., Clarke, C., & Rafiei, D. (2023). Evaluating Open-Domain Question Answering in the Era of Large Language Models ArXiv, abs/2305.06984. https://doi.org/10.48550/arXiv.2305.06984
Bayat, F. F., Qian, K., Han, B., Sang, Y., Belyi, A., Khorshidi, S., … Li, Y. (2023). FLEEK: Factual Error Detection and Correction With Evidence Retrieved From External Knowledge ArXiv, abs/2310.17119. https://doi.org/10.48550/ARXIV.2310.17119
Zhong, W., Xie, Y., & Lin, J. (2023). Answer Retrieval for Math Questions Using Structural and Dense Retrieval Presented at the Retrieval for Math Questions Using Structural and Dense Retrieval Primary Tabs View conference. https://doi.org/10.1007/978-3-031-42448-9_18
Zong, S., Seltzer, J., Pan, J., Cheng, K., & Lin, J. (2023). Which Model Shall I Choose? Cost/Quality Trade-Offs for Text Classification Tasks ArXiv, abs/2301.07006. https://doi.org/10.48550/arXiv.2301.07006
Thakur, N., Ni, J., Abrego, G. H. andez \, Wieting, J., Lin, J., & Cer, D. (2023). Leveraging LLMs for Synthesizing Training Data Across Many Languages In Multilingual Dense Retrieval ArXiv, abs/2311.05800. https://doi.org/10.48550/ARXIV.2311.05800