Umeå universitets logga

umu.sePublikationer
Ändra sökning
Länk till posten
Permanent länk

Direktlänk
Ågren, Ola M
Alternativa namn
Publikationer (9 of 9) Visa alla publikationer
Ågren, O. (2015). AMBiDDS: A system for Automatic Mining of BIg Discrete Data-Sets. In: 2015 INTERNATIONAL CONFERENCE ON COMPUTATIONAL SCIENCE AND COMPUTATIONAL INTELLIGENCE (CSCI): . Paper presented at International Conference on Computational Science and Computational Intelligence (CSCI), DEC 07-09, 2015, Las Vegas, NV (pp. 424-427).
Öppna denna publikation i ny flik eller fönster >>AMBiDDS: A system for Automatic Mining of BIg Discrete Data-Sets
2015 (Engelska)Ingår i: 2015 INTERNATIONAL CONFERENCE ON COMPUTATIONAL SCIENCE AND COMPUTATIONAL INTELLIGENCE (CSCI), 2015, s. 424-427Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

This paper introduces an automatic algorithm that can be seen as an extension to the Eclat algorithm, as well as a corresponding proof of concept prototype. It uses inverted indices and statistical pruning of the possible solution space as early as possible.

Nyckelord
Data mining, inverted indices, statistical pruning
Nationell ämneskategori
Beräkningsmatematik
Identifikatorer
urn:nbn:se:umu:diva-124694 (URN)10.1109/CSCI.2015.142 (DOI)000380405100076 ()2-s2.0-84964413106 (Scopus ID)978-1-4673-9795-7 (ISBN)
Konferens
International Conference on Computational Science and Computational Intelligence (CSCI), DEC 07-09, 2015, Las Vegas, NV
Tillgänglig från: 2016-10-28 Skapad: 2016-08-22 Senast uppdaterad: 2023-03-24Bibliografiskt granskad
Ågren, O. M. (2015). Student-graded oral presentations. International Journal of Engineering Pedagogy, 5(4), 76-78
Öppna denna publikation i ny flik eller fönster >>Student-graded oral presentations
2015 (Engelska)Ingår i: International Journal of Engineering Pedagogy, ISSN 2192-4880, Vol. 5, nr 4, s. 76-78Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

We describe a way to use peer-graded oral presentations as a way of reducing the load on the teacher, and show that almost identical results as can be achieved as with teacher graded presentations. Moreover, we have found that very little in the form of explicit criteria are needed.

Ort, förlag, år, upplaga, sidor
Kassel University Press GmbH, 2015
Nyckelord
Didactics, Peer assessment, Teacher offloading
Nationell ämneskategori
Utbildningsvetenskap
Identifikatorer
urn:nbn:se:umu:diva-114390 (URN)10.3991/ijep.v5i4.4841 (DOI)000366993800009 ()
Tillgänglig från: 2016-01-18 Skapad: 2016-01-18 Senast uppdaterad: 2018-06-07Bibliografiskt granskad
Ågren, O. M. (2012). The ProT Nordic Web Dataset. In: Hamid R. Arabnia, Victor A. Clincy, Leonidas Deligiannidis, Andy Marsh, Ashu M. G. Solo (Ed.), Proceedings of the International Conference on Internet Computing: ICOMP 2012. Paper presented at The 2012 International Conference on Internet Computing (ICOMP'12) (pp. 125-128). Las Vegas, Nevada: CSREA Press
Öppna denna publikation i ny flik eller fönster >>The ProT Nordic Web Dataset
2012 (Engelska)Ingår i: Proceedings of the International Conference on Internet Computing: ICOMP 2012 / [ed] Hamid R. Arabnia, Victor A. Clincy, Leonidas Deligiannidis, Andy Marsh, Ashu M. G. Solo, Las Vegas, Nevada: CSREA Press, 2012, s. 125-128Konferensbidrag, Enbart muntlig presentation (Refereegranskat)
Abstract [en]

In this paper we present a free dataset, usable for testing web search engines.  The dataset corresponds to a snapshot of the Nordic part of the Internet in early 2007 and is highly abstracted, with numbers representing each web page.  The released dataset consists of three parts; a graph, 76 sets of pages containing each tested word combination, and some files to use when calculating relevance of the resulting sets of algorithms/search engines. We also present a new compound statistic as well as statistical results for some search engine and information retrieval algorithms.

Ort, förlag, år, upplaga, sidor
Las Vegas, Nevada: CSREA Press, 2012
Nyckelord
Nordic Web Dataset, Search Engine Evaluation, Relevance Metrics
Nationell ämneskategori
Datavetenskap (datalogi)
Forskningsämne
datalogi
Identifikatorer
urn:nbn:se:umu:diva-64012 (URN)1-60132-220-8 (ISBN)
Konferens
The 2012 International Conference on Internet Computing (ICOMP'12)
Projekt
ProT
Tillgänglig från: 2013-01-11 Skapad: 2013-01-11 Senast uppdaterad: 2018-06-08Bibliografiskt granskad
Ågren, O. (2011). Using the ProT Nordic Web Dataset.
Öppna denna publikation i ny flik eller fönster >>Using the ProT Nordic Web Dataset
2011 (Engelska)Rapport (Övrigt vetenskapligt)
Abstract [en]

In this paper we present a free dataset, usable for testing web search engines. The dataset corresponds to a snapshot of the Nordic part of the Internet back in early 2007 and is highly abstracted, with numbers representing each web page. The released dataset consists of three parts; a graph, 76 sets of pages containing each tested word combination, and some files to use when calculating relevance of the resulting sets of algorithms/search engines. We also present statistics for some search engine algorithms.

Förlag
s. 29
Serie
Report / UMINF, ISSN 0348-0542 ; 13
Nyckelord
Nordic Web Dataset, Search Engine Evaluation, Relevance Metrics
Nationell ämneskategori
Annan elektroteknik och elektronik
Forskningsämne
data- och systemvetenskap
Identifikatorer
urn:nbn:se:umu:diva-49307 (URN)
Projekt
ProT
Tillgänglig från: 2011-11-10 Skapad: 2011-11-07 Senast uppdaterad: 2018-06-08Bibliografiskt granskad
Ågren, O. (2008). S²ProT: Rank Allocation by Superpositioned Propagation of Topic-Relevance. International Journal of Web Information Systems, 4(4), 416-440
Öppna denna publikation i ny flik eller fönster >>S²ProT: Rank Allocation by Superpositioned Propagation of Topic-Relevance
2008 (Engelska)Ingår i: International Journal of Web Information Systems, ISSN 1744-0084, Vol. 4, nr 4, s. 416-440Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

Purpose – The purpose of this paper is to assign topic-specific ratings to web pages.

Design/methodology/approach – The paper uses power iteration to assign topic-specific rating values (called relevance) to web pages, creating a ranking or partial order among these pages for each topic. This approach depends on a set of pages that are initially assumed to be relevant for a specific topic; the spatial link structure of the web pages; and a net-specific decay factor designated ξ.

Findings – The paper finds that this approach exhibits desirable properties such as fast convergence, stability and yields relevant answer sets. The first property will be shown using theoretical proofs, while the others are evaluated through stability experiments and assessments of real world data in comparison with already established algorithms.

Research limitations/implications – In the assessment, all pages that a web spider was able to find in the Nordic countries were used. It is also important to note that entities that use domains outside the Nordic countries (e.g..com or.org) are not present in the paper's datasets even though they reside logically within one or more of the Nordic countries. This is quite a large dataset, but still small in comparison with the entire worldwide web. Moreover, the execution speed of some of the algorithms unfortunately prohibited the use of a large test dataset in the stability tests.

Practical implications – It is not only possible, but also reasonable, to perform ranking of web pages without using Markov chain approaches. This means that the work of generating answer sets for complex questions could (at least in theory) be divided into smaller parts that are later summed up to give the final answer.

Originality/value – This paper contributes to the research on internet search engines.

Nyckelord
Information retrieval, Search engines, Spatial data structures, Worldwide web
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
urn:nbn:se:umu:diva-11236 (URN)10.1108/17440080810919477 (DOI)2-s2.0-84886421149 (Scopus ID)
Tillgänglig från: 2008-12-01 Skapad: 2008-12-01 Senast uppdaterad: 2023-03-24Bibliografiskt granskad
Ågren, O. (2006). Assessment of WWW-Based Ranking Systems for Smaller Web Sites. INFOCOMP Journal of Computer Science, 5(2), 45-55
Öppna denna publikation i ny flik eller fönster >>Assessment of WWW-Based Ranking Systems for Smaller Web Sites
2006 (Engelska)Ingår i: INFOCOMP Journal of Computer Science, ISSN 1807-4545, Vol. 5, nr 2, s. 45-55Artikel i tidskrift (Refereegranskat) Published
Abstract [en]

A comparison between a number of search engines from three different families (HITS, PageRank, and Propagation of Trust) is presented for a small web server with respect to perceived relevance. A total of 307 individual tests have been done and the results from these were disseminated to the algorithms, and then handled using confidence intervals, Kolmogorov-Smirnov and ANOVA. We show that the results can be grouped according to algorithm family, and also that the algorithms (or at least families) can be partially ordered in order of relevance.

Nyckelord
Assessment, Search engines, HITS, PageRank, Propagation of Trust, eigenvectors
Nationell ämneskategori
Datavetenskap (datalogi)
Identifikatorer
urn:nbn:se:umu:diva-8415 (URN)
Tillgänglig från: 2008-01-21 Skapad: 2008-01-21 Senast uppdaterad: 2018-06-09Bibliografiskt granskad
Ågren, O. (2003). CHiC: A Fast Concept Hierarchy Constructor for Discrete or Mixed Mode Databases. In: SEKE 2003: Proceedings of the Fifteenth International Conference on Software Engineering & Knowledge Engineering. Paper presented at The 15th International Conference on Software Engineering and Knowledge Engineering (SEKE'03), San Fransisco, July 1-3, 2003 (pp. 250-258). Knowledge Systems Institute
Öppna denna publikation i ny flik eller fönster >>CHiC: A Fast Concept Hierarchy Constructor for Discrete or Mixed Mode Databases
2003 (Engelska)Ingår i: SEKE 2003: Proceedings of the Fifteenth International Conference on Software Engineering & Knowledge Engineering, Knowledge Systems Institute, 2003, s. 250-258Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

In this paper we propose an algorithm that automatically creates concept hierarchies or lattices for discrete databases and datasets. The reason for doing this is to accommodate later data mining operations on the same sets of data without having an expert create these hierarchies by hand.

Each step of the algorithm will be examined; We will show inputs and output for each step using a small example. The theoretical upper bound of the complexity for each part of the algorithm will be presented, as well as real time measurements for a number of databases. We will finally present a time model of the algorithm in terms of a number of attributes of the databases

Ort, förlag, år, upplaga, sidor
Knowledge Systems Institute, 2003
Nyckelord
Data Mining, Data Preprocessing, Hierarchy Generation, Lattice Generation
Nationell ämneskategori
Datavetenskap (datalogi)
Forskningsämne
administrativ databehandling
Identifikatorer
urn:nbn:se:umu:diva-22348 (URN)1891706128 (ISBN)
Konferens
The 15th International Conference on Software Engineering and Knowledge Engineering (SEKE'03), San Fransisco, July 1-3, 2003
Projekt
CHiC
Tillgänglig från: 2009-05-06 Skapad: 2009-05-06 Senast uppdaterad: 2018-06-08Bibliografiskt granskad
Ågren, O. (2002). Automatic Generation of Concept Hierarchies for a Discrete Data Mining System. In: Hamid R. Arabnia, Youngsong Mun, Bhanu Prasad (Ed.), International Conference on Information and Knowledge Engineering (IKE '02): . Paper presented at The 2002 International Conference on Information and Knowledge Engineering (IKE '02), June 24-27, 2002, Las Vegas, USA (pp. 287-293). CSREA Press
Öppna denna publikation i ny flik eller fönster >>Automatic Generation of Concept Hierarchies for a Discrete Data Mining System
2002 (Engelska)Ingår i: International Conference on Information and Knowledge Engineering (IKE '02) / [ed] Hamid R. Arabnia, Youngsong Mun, Bhanu Prasad, CSREA Press, 2002, s. 287-293Konferensbidrag, Publicerat paper (Refereegranskat)
Abstract [en]

In this paper we propose an algorithm for automatic creation of concept hierarchies from discrete databases and datasets. The reason for doing this is to accommodate later data mining operations on the same set of data without having an expert create these hierachies by hand.

We will go through the algorithm thoroughly and show the results from each step of the algorithm using a (small) example. We will also give actual execution times for our prototype for non-trivial example data sets and estimates of the complexity of the algorithm in terms of the number of records and the number of distinct data values in the data set.

Ort, förlag, år, upplaga, sidor
CSREA Press, 2002
Nyckelord
Data Mining, Data Preprocessing, Hierarchy Generation
Nationell ämneskategori
Datavetenskap (datalogi)
Forskningsämne
administrativ databehandling
Identifikatorer
urn:nbn:se:umu:diva-22344 (URN)1892512971 (ISBN)
Konferens
The 2002 International Conference on Information and Knowledge Engineering (IKE '02), June 24-27, 2002, Las Vegas, USA
Projekt
CHiC
Tillgänglig från: 2009-05-06 Skapad: 2009-05-06 Senast uppdaterad: 2018-06-08Bibliografiskt granskad
Ågren, O. (2001). AlgExt: an Algorithm Extractor for C Programs. Umeå: Department of Computing Science, Umeå University
Öppna denna publikation i ny flik eller fönster >>AlgExt: an Algorithm Extractor for C Programs
2001 (Engelska)Rapport (Övrigt vetenskapligt)
Abstract [en]

ALGEXT is a program that extracts strategic/block comments from C source files to improve maintainability and to keep documentation consistent with source code. This is done by writing the comments in the source code in what we call extractable algorithms, describing the algorithm used in the functions.

ALGEXT recognizes different kinds of comments:

  • Strategic comments are comments that proceed a block of code, with only whitespace preceding it on the line,
  • Tactical comments are comments that describes the code that precedes it on the same line,
  • Function comments are comments immediately preceding a function definition, describing the function,
  • File comments are comments at the head of the file, before any declarations of functions and variables, and finally
  • Global comments are comments within the global scope, but not associated with a function.

Only strategic comment are used as basis for algorithm extraction in ALGEXT.

The paper discusses the rationale for ALGEXT and describes its implementation and usage. Examples are presented for clarification of what can be done with ALGEXT.

Our experience shows that students who use ALGEXT for preparing theirassignments tend to write about 66% more comments than non-ALGEXT users.

Ort, förlag, år, upplaga, sidor
Umeå: Department of Computing Science, Umeå University, 2001. s. 15
Serie
Report / UMINF, ISSN 0348-0542 ; 2001:11
Nyckelord
Extractable algorithms, Embedded information, C
Nationell ämneskategori
Datavetenskap (datalogi)
Forskningsämne
administrativ databehandling
Identifikatorer
urn:nbn:se:umu:diva-22350 (URN)
Distributör:
Institutionen för datavetenskap, 90187, Umeå
Projekt
AlgExt
Tillgänglig från: 2009-05-06 Skapad: 2009-05-06 Senast uppdaterad: 2018-06-08Bibliografiskt granskad
Organisationer

Sök vidare i DiVA

Visa alla publikationer