- Research Article
16
- 10.1016/j.procs.2015.03.122
A Novel Approach to HTML Page Creation Using Neural Network
- Jan 01, 2015
- Procedia Computer Science
- Aparna Halbe + 1 more +1
A Novel Approach to HTML Page Creation Using Neural Network
In the brief history of the World Wide Web (WWW), much has changed. Millions of web pages have been published in a relatively short time. Next to the Web content, the one of the most dynamic aspects of the WWW is the development of HyperText Markup Language (HTML). This paper explores the various versions of HTML and gives a status report on HTML standards development. A discussion of possible future trends is also included.
A Novel Approach to HTML Page Creation Using Neural Network
A Novel Approach to HTML Page Creation Using Neural Network
유사 패턴을 갖는 HTML 문서의 XML 자동 변환
최근 들어, WWW(World Wide Web)의 급속한 보급으로 많은 양의 정보가 생성되고 있다. 이로 인하여 웹은 이제 정보 교환의 도구로서 뿐 아니라 정보의 저장소로 인식되게 되었다. 현재 웹상의 많은 문서들은 HTML(Hypertext Markup Language)을 사용하여 제작되었다. HTML은 간단하고 배우기가 쉬운 반면, 데이터에 대한 기술을 명확하게 하지 못하는 단점으로 인해 정보 검색에 있어서 효율성을 제공하지 못한다. 이를 보완하기 위한 방법 중에 하나가 구조적인 언어로 부상하고 있는 XML(eXtensible Markup Language) 문서로 변환하는 것이다. XML은 웹 상에서 데이터 교환을 위해 제안된 표준 메타 언어이다. 효과적인 데이터의 교환을 위해, XML은 DTD(Document Type Definition)를 통하여 문서의 구조를 기술할 수 있고 사용자가 원하는 대로 정의할 수 있다. 이러한 구조적 유동성은 웹에서 운용되는 모든 데이터를 통합, 저장, 처리할 수 있는 기반을 제공한다. 본 논문에서는 특히 유사한 패턴을 갖는 HTML 문서의 구조를 분석하고 그에 관련된 경로 정보를 인식하는 방식을 이용하여 XML 문서로의 변환을 자동적으로 수행할 수 있는 XML 변환기를 구현하였다. Recently, WWW(World Wide Web) has become a source of a large amount of information, and is now recognized not only as an information-sharing tool, but also as an information repository. Currently, the majority of documents on the web were created using HTML(Hypertext Markup Language). Although HTML is simple and easy to learn, its inherent lack of describing document structure makes it difficult to retrieve information effectively. One possible solution would be to convert such HTML documents into XML (extensible Markup Language) documents. This is a standard markup language for exchanging data on the web. It can describe a document structure freely by defining its own DTD (Document Type Definition). This makes it possible to integrate, store, and retrieve data on the web efficiently In this paper, we will propose a converter that automatically converts HTML documents with similar pattern into XML documents by analyzing the document structure and recognizing its path information.
Read moreAnalysis of the HTML to XML Conversion Method
In this paper, the features of HTML and XML were compared, and expounds the necessity of transition from HTML to XML, and finally introduces document conversion, XHTML conversion and intelligent instead of three kinds of HTML to XML conversion method. Keyword: HTML;XML;conversion method;XHTML Characteristics of HTML HTML (Hypertext Markup Language) is a universal language for creating Webpage and to release information, it is YISHION text stored in the form of organization, to label definition document. Provided the cross plat form file sharing. In the HTML document, can be embedded in other objects, such as electronic forms, video, audio and various applications and so on, through the uniform resource locator can realize the hypertext links between Web nodes. HTML has the following features, format and grammar is relatively simple, easy to learn, will the data with some control markers, even if no programming experience can easily use HTML to design Webpage; and all of the control tag HTML is fixed [1], the number is limited, provides functions and related properties the setting is fixed, easy to remember; rules more flexible, such as the control of English marker size marker to write in no difference. In addition, the control flag must have the end tag corresponding to no strict requirements. Need the simplicity of HTML is more suitable for low cost information publishing; HTML as Web common information description method of strong universality, can realize different platform document sharing; to create more flexible, HTML document is a plain text file, can use a variety of editing tools for creating. The major disadvantage of HTML is: (1) Performance is too simple. (2) Link easily broken, chain destination address changes, chain source cannot automatically correct. (3) The flowers during the retrieval time are longer, the retrieved content targeted poor, and many returned results. (4) Poor scalability, HTML tag set is fixed, not to allow users to define their International Symposium on Computers & Informatics (ISCI 2015) © 2015. The authors Published by Atlantis Press 64 own identity. (5) The lack of semantics, HTML is a marker, it will not reveal the nature of the information content, and the computer cannot know the exact meaning of each section of text. HTML and XML are used for the network information organization and communication, are in the form of texts written in storage, and are structured information based on international standards [2]. But HTML only shows data seems to be what kind of, and the XML is that data is what mean. Using XML can create their own tags, these tags can more accurately describe the user what they want, but HTML can not be used to define a new application, and this is the biggest difference between XML and HTML. XML has overcome some limitations of the HTML, has the broad application prospect. The necessity of transformation Most of the current Webpage still is mainly composed of HTML. HTML of the inherent shortcomings of the original network information organization mode can't meet the requirements of the development of the new. Because the relevant development of the XML technology continues to mature, more and more websites gradually in XML design. In this process, not only will the new content is stored in XML format and transmission, but also consider compatible with the original data, so it is necessary to convert existing HTML Webpage into XML data form more flexible processing and application. At the same time, to the previous accumulated HTML documents continue to play a role in the new environment, the XML conversion is undoubtedly a solution. HTML to XML conversion is helpful for Webpage information integration, extraction, retrieval, filtering or mining analysis. (1) Facilitate the information integration of heterogeneous system under Network Environment In the network environment, because of the existence of the system platform and database operation of heterogeneous, leading to the exchange of information and work with difficulty. One of the data exchange between heterogeneous systems is to adopt a unified information exchange format. XML because of its custom and scalability advantages [3], which is convenient for expressing various types of data between heterogeneous databases, and can be used as middleware, unify the data interface, convenient for the exchange of information between different databases and colleagues. XML can be used to construct the data layer integration, will transform the source data into the data integration, simplification of integration system query translation, query mechanisms provide unified multi data source for the user, in a unified way using a variety of data from different sources of data, different shielding each data source in the structure, running environment on the. (2) The organization and management for network information
Read moreArchive of sidescan-sonar data and DGPS navigation data, collected during USGS cruise SEAX 96004, New York Bight, 1 May-9 June, 1996
This DVD-ROM contains digital high resolution sidescan-sonar data collected during USGS cruise SEAX 96004 aboard the R/V Seaward Explorer. The coverage lies along New York Bight. This DVD-ROM (Digital Versatile Disc-Read Only Memory) has been produced in accordance with the UDF DVD- ROM Standard and is therefore capable of being read on any computing platform that has appropriate DVD-ROM driver software installed. Access to the data and information contained on this DVD-ROM was developed using the HyperText Markup Language (HTML) utilized by the World Wide Web (WWW) project. Development of the DVD-ROM documentation and user interface in HTML allows a user to access the information by using a variety of WWW information browsers (i.e. NCSA Mosaic, Netscape) to facilitate browsing and locating information and data. To access the information contained on this disc with a WWW client browser, open the file 'index.htm' at the top level directory of this DVD-ROM with your selected browser. The HTML documentation is written utilizing some HTML 4.0 enhancements. The disc should be viewable by all WWW browsers but may not properly format on some older WWW browsers. Also, some links to USGS collaborators are available on this DVD-ROM. These links are only accessible if access to the Internet is available during browsing of the DVD-ROM. Software is available on this DVD-ROM for viewing and processing the individual swaths using computer systems running the UNIX operating system.
Read moreCore Semantics for Public Ontologies
: The World Wide Web contains a large amount of information which is currently being represented using the Hypertext Markup Language (HTML) and the Extensible Markup Language (XML). However, HTML and XML have a limited capability to describe the relationships (schemas or ontologies) with respect to objects. The DARPA Agent Markup Language (DAML) through the use of ontologies provides a very powerful way to describe objects and their relationships to other objects. DAML is an extension to XML and the Resource Description Framework (RDF). This effort focused on development of semantic transfer protocols for exchange of information and communication between semantically competent agents. It also developed an extended theory semantically expressive description languages used by and shared between agents.
Read moreHypertext Markup Language, the Wireless Way
As discussed in chapter 4, “The Wireless World Wide Web,” the HyperText Markup Language (HTML) has significant advantages as a medium for wireless access terminals. However, you must approach its use with the constraints of handheld devices in mind If you are used to developing in HTML for desktop browsers, you will have to start thinking a little differently, making sure you stick to HTML tags appropriate for the wireless world. In this chapter, I first review the different versions of HTML in light of their suitability for use with the wireless Web. I then walk through a tag-by-tag discussion of using HTML in your development of wireless content.
Read moreGIS-based mapping potential sites for micro-hydro power plants in West Sumatera
The increased energy demand, global warming and other environmental problems caused by the use of fossil fuels have brought severe challenges to West Sumatera, it's make the government Province has set targets to improve the share of renewable energy in the energy structure and promote the construction of hydropower plants. In this research, we are develops geographic information system (GIS) based clustering. Data was collected from open map-servers and geocoded by open data kit package and data geocoding tools. The Web-based system is designed used a program language of Preprocessor Hypertext (PHP) and Hypertext Markup Language (HTML) which connected to PostGIS database to store the data. The system provides Web-based GIS informed about the location of potential micro hydro power plant each village in West Sumatera and analyst tool for pattern detection through K-means clustering. The result of this research, through which end users can detect micro-hydro power plant distribution pattern based cluster selecting area, and spatial test parameters, which can be used to know the consumer energy usability in the province of West Sumatera.
Read moreOracle Application Express
The World Wide Web (WWW) was invented by renowned British computer scientist Sir Tim Berners-Lee in 1989. He had proposed three fundamental concepts that have today become ubiquitous to many. It provides us timely information, a way to access services, and more. These principles include the HyperText Markup Language (HTML), the Uniform Resource Identifier (URI) (also commonly known as URL), and the Hypertext Transfer Protocol (HTTP). URIs are human-friendly addresses on the Internet that allow us to identify and locate web resources, which are then transmitted over the network, to web browsers, using HTTP. The most common web resources are web pages formatted using HTML.
Read moreSecuring e-commerce against SQL injection, cross site scripting and broken authentication
World Wide Web (WWW) has been introduced in 1980s and is widely been used until today. With WWW service, publisher able to host a website in form of hypertext using Hypertext Mark-up Language (HTML). In addition, Cascading Stylesheet (CSS) is always used with HTML to manage the layout of the webpage. Over the years, the capability of HTML and CSS is getting enhanced to create a more responsive webpage. However, all these webpages creation is more towards information sharing and does not really handle user inputs. Hence, in this project, the security measures are proposed to counter these threats will be compiled as a library to be usable in any PHP-based web application. A basic but fully functional e-commerce application is developed for the testing of the proposed security features to countermeasures the mentioned vulnerabilities.
Read moreWWW (World Wide Web) Communication and Publishing of Structural Formulas by XyMML (XyM Markup Language)
A tool for displaying and communicating chemical structural formulas has been developed on the basis of XyMML (XyM Markup Language), where a XyMML document according to the XML (Extensible Markup Language) specification has been transformed into an HTML (HyperText Markup Language) document by means of a translator program due to XSLT (Extensible Stylesheet Language Transformations). During this process, XyMML data written in such a XyMML document have been converted into XyM notations embedded in such an HTML document, which is browsed by virtue of a World Wide Web (WWW) browser including the XyMJava system. Another tool for printing chemical structural formulas has been developed so that the same XyMML document has been transformed into a XyMTeX document by means of XSLT. The resulting XyMTeX document has been used to print a document containing structural formulas through the TeX/LaTeX typesetting system. Thereby, the XyMML and the related techniques have been shown to have the potentiality of serving as a kernel for integrating WWW communication, electronic publishing, and conventional publishing in chemistry.
Read moreDesign of on-line Interactive Data Acquisition and Control System for embedded real time applications
Design of on-line embedded web server is a challenging part of many embedded and real time data acquisition and control system applications. The World Wide Web is a global system of interconnected computer networks that use the standard Internet Protocol Suite (TCP/IP) to serve billion of users worldwide and allows the user to interface many real time embedded applications like data acquisition, Industrial automations and safety measures etc,. This paper approached towards the design and development of on-line Interactive Data Acquisition and Control System (IDACS) using ARM based embedded web server. It can be a network, intelligent and digital distributed control system. Single chip IDACS method improves the processing capability of a system and overcomes the problem of poor real time and reliability. This system uses ARM9 Processor portability with Real Time Linux operating system (RTLinux RTOS) it makes the system more real time and handling various processes based on multi tasking and reliable scheduling mechanisms. Web server application is ported into an ARM processor using embedded `C' language. Web pages are written by Hyper text markup language (HTML); it is beneficial for real time IDACS, Mission critical applications, ATM networks and more.
Read moreA modelling methodology for the HTML knowledge base
The hypertext markup language (HTML) has taken an irreplaceable position in the field of the internet. In this study, the complicated text-based HTML syntax is modelled as a shareable and computable knowledge base (KB). The proposed modelling methodology is sound and complete because it covers all constraints and semantics defined in the HTML. The result of the modelling methodology is a KB with a class hierarchy, constraints, and inference rules. It is programming language independent that can be used by any language with access interfaces to a KB. Based on the HTML KB, one can develop various tools and intelligent systems with all the advantages that KB system is promised to provide. For demonstration purpose, an expert system, the web content accessibility guidelines (WCAG) validation system, was built. It validates whether a web page meets the WCAG regulation and the HTML KB serves as a shareable KB in the system.
Read moreDeus ex Machina: Navigating between the Lines
GOOD EVENING. Tonight I've been asked to talk to you about HTML – hypertext markup language – and its performative characteristics; its multimedia capacity; its non-linear structure; its interactive possibilities; its real-time relationship with its readers slash navigators slash audience; and its potential interest to writers cum artists cum performers.
Read moreProgramming in the Life Sciences #11: HTML
HTML (HyperText Markup Language), the language of the web, is no longer the only language of the web. But it still is the primary language in which source code of webpages is shared. Originally, HTML pages were always static: the only HTML source of a web page was that was downloaded from a website. Nowadays, much HTML the is visualized in your web browser, is generated on the fly with JavaScript.
Read moreWeb Services Management
The widespread use and expansion of the World Wide Web has revolutionized the discovery, access, and retrieval of information. The Internet has become the doorway to a vast information base and has leveraged the access to information through standard protocols and technologies like HyperText Markup Language (HTML), active server pages (ASP), Java server pages (JSP), Web databases, and Web services. Web services are software applications that are accessible over the World Wide Web through standard communication protocols. A Web service typically has a Webaccessible interface for its clients at the front end, and is connected to a database system and other related application suites at the back end. Thus, Web services can render efficient Web access to an information base in a secured and selective manner. The true success of this technology, however, largely depends on the efficient management of the various components forming the backbone of a Web service system. This chapter presents an overview and the state of the art of various management approaches, models, and architectures for Web services systems toward achieving quality of service (QoS) in Web data access. Finally, it discusses the importance of autonomic or self-managing systems and provides an outline of our current research on autonomic Web services.
Read more