This site requires Cookies enabled in your browser for login.
Updating ...
WaterNet Home
WaterNet
for
pour le
Canada
Menu
WaterNet
Home
GWFO
Home
Catalogue
Master Index
Data
Centre
X
Find Data By Variable Find Data By Site, Facility, or Deployable Show Near-realtime Telemetry (7 day)
Collections
X
Defaults
Select All
Websites
X
Global Water Futures Observatories (GWFO) Global Water Futures (GWF) Global Institute for Water Security (GIWS) International Network of Alpine Research Catchment Hydrology
Legacy Research Programs
X
Changing Cold Regions Network (CCRN) Drought Research Initiative (DRI) International Network of Alpine Research Catchment Hydrology (Legacy Site) Improving Processes & Parameterization for Prediction in Cold Regions Hydrology (IP3) The Mackenzie Global Energy and Water Cycle Experiment (GEWEX) Study (MAGS)
Legacy sites
Map
Utilities
X
Account Settings Create a New Record Record List Alias List Editor
Edit Data Centre
Data Types
. . .
X
Clear
Select All
Advanced Search
Go to Top⇡
Related items loading ...
Fetching Chart ...
Publication Additional Information Download
Publication Type
Journal Article
Authorship
Al-Omari, F., Roy, C. K., & Chen, T.
Title
SemanticCloneBench: A Semantic Code Clone Benchmark using Crowd-Source Knowledge
Year
2020
Publication Outlet
In 2020 IEEE 14th International Workshop on Software Clones (IWSC) (pp. 57-63). IEEE
DOI
https://doi.org/10.1109/IWSC50091.2020.9047643
Citation
Al-Omari, F., Roy, C. K., & Chen, T. (2020). SemanticCloneBench: A Semantic Code Clone Benchmark using Crowd-Source Knowledge. In 2020 IEEE 14th International Workshop on Software Clones (IWSC) (pp. 57-63). IEEE. https://doi.org/10.1109/IWSC50091.2020.9047643
Abstract
Not only do newly proposed code clone detection techniques, but existing techniques and tools also need to be evaluated and compared. This evaluation process could be done by assessing the reported clones manually or by using benchmarks. The main limitations of available benchmarks include: they are restricted to one programming language; they have a limited number of clone pairs that are confined within the selected system(s); they require manual validation; they do not support all types of code clones. To overcome these limitations, we proposed a methodology to generate a wide range of semantic clone benchmark(s) for different programming languages with minimal human validation. Our technique is based on the knowledge provided by developers who participate in the crowd-sourced information website, Stack Overflow. We applied automatic filtering, selection and validation to the source code in Stack Overflow answers. Finally, we build a semantic code clone benchmark of 4000 clones pairs for the languages Java, C, C# and Python.
Program Affiliations
GWF: Global Water Futures
Publication Stage
Published
Download Links
https://doi.org/10.1109/IWSC50091.2020.9047643
© 2026 - WaterNet Version 2026-07-24
Global Water Futures Observatories
Powered by
G W F Net
T-2022-12-05-o1M6M6o2FiaUK5h9kC2MeAo3Q Publication 1.0